12 WHITEPAPERS 4 RESEARCH PAPERS BUILDING SINCE 2011
RESEARCH — WHITE PAPERS

Every project starts as a question.

Technical white papers spanning work from 2018 to today, plus ongoing academic research in mechanistic interpretability and on-device AI. Grouped by the year each project was actually built, not when we got around to documenting it.

2026 10 papers
AUG 2026 PAPER v1.0

Airwave

A native iPadOS app that turns a tablet and RTL‑SDR dongle into a full‑spectrum scanner and automatic emitter classifier — 24 MHz to 1.766 GHz, with a Metal waterfall display and DSP pipeline in Accelerate/vDSP.

RTL-SDRAccelerate / vDSPMetalBLEiPadOS
Read PDF
AUG 2026 PAPER v1.0

MLX Serving Gateway

A production‑grade Swift 6 inference server on Hummingbird 2 and MLX Swift — OpenAI‑compatible API, prefix‑trie KV cache, deadline‑bounded batch assembler, and LRU model pool. No Python, no Docker.

Swift 6Hummingbird 2KV CacheOpenAI APIApple Silicon
Read PDF
AUG 2026 PAPER v1.0

ScanForge

GPU‑accelerated TSDF volumetric reconstruction for watertight 3D scanning on consumer LiDAR hardware — real‑time depth‑frame integration via Metal compute, Marching Cubes surface extraction, OBJ/STL export ready for any slicer.

TSDFMetal ComputeLiDARMarching Cubes3D Print
Read PDF
AUG 2026 PAPER v1.0

Scholar

On‑device neural TTS for academic reading on iPad — Kokoro 82M (StyleTTS2) reimplemented in MLX Swift, with a full G2P frontend, PLBERT prosody encoder, and iSTFT vocoder. Zero network calls at inference time.

Kokoro / StyleTTS2MLX SwiftG2POn-DeviceiPadOS
Read PDF
JUL 2026 PAPER v1.0

SDSTK Studio

A native Swift canvas for visual data science and compound model orchestration on Apple Silicon.

MLX-NativeVisual ProgrammingMixture of ExpertsMCPiPad / macOS
Read PDF
JUL 2026 PAPER v1.0

QuantForge

A native macOS and iPadOS app that brings LLM quantization — GGUF and MLX pipelines — out of the terminal and into a direct‑manipulation interface. Inspect architecture, choose a format, watch live progress, get before/after benchmarks.

GGUFMLXQuantizationllama.cppmacOS / iPadOS
Read PDF
JUN 2026 PAPER v1.0

mlxMesh / Open Inference Mesh

A federated protocol for privacy‑tiered, measured‑accountability AI compute across wide‑area networks.

WAN FederationMoE ShardingEd25519Secure EnclaveOpen Protocol
Read PDF
MAY 2026 PAPER v1.0

ExoControlCenter

A native iPad control plane for distributed AI inference clusters built on exo.

exoiPad NativePipeline ParallelCluster OpsMLX
Read PDF
APR 2026 PAPER v1.0

Model Cartography

A universal platform for neural network interpretability, attribution, and surgical intervention.

Mechanistic Interp.SAEMoE RoutingAI ActCross-Platform
Read PDF
MAR 2026 PAPER v1.0

SwiftSci

An MLX‑native Swift reimplementation of the Python scientific data science stack.

MLXSwiftAccelerate / LAPACKNo GILApple Silicon
Read PDF
2020 1 paper · updated 2026
FEB 2020 ↻ UPD 2026 PAPER v1.0

Censor

On‑device few‑shot annotation and dataset organization for Apple platforms. Originally built 2020, actively maintained since.

Few-ShotOn-DeviceVision FrameworkiPad / macOSZero Network Calls
Read PDF
2018 1 paper · where it started
2018 · ORIGIN PAPER v1.0

ModelBuilder

A universal platform for on‑device machine learning training, inference, and model management — letting anyone build Core ML models from their own data. Our earliest work.

Core MLCreate MLSwiftUIiPad / macOSOn-Device
Read PDF
Academia — Independent Research

Research we publish in the open.

Mechanistic interpretability, sparse autoencoders, and on-device AI efficiency — academic papers in active review at ICLR and NeurIPS. Preprints available now.

2026 4 papers · in review
ICLR 2026 WORKING PAPER

Circuit Tracing at Scale

Circuit‑level mechanisms identified in small models generalize to larger models with additional redundancy structure. Targeted path‑patching combined with sparse decomposition on Apple Silicon — IOI circuits measured on Llama and Pythia.

Mechanistic Interp.Path PatchingIOIApple Silicon
Abstract & PDF
NeurIPS 2026 WORKING PAPER

Sparse Autoencoders as Feature Finders

TopK SAE training dynamics across Llama, Mistral, and Qwen: dictionary collapse, dense‑feature degeneracy, and held‑out reconstruction. Evaluation metric choice determines which architecture appears superior — a hidden confound in prior work.

SAETopKDictionary LearningLlama / Mistral / Qwen
Abstract & PDF
ICLR 2026 WORKING PAPER

Distributed Interpretability

Activation streaming across two nodes via Thunderbolt: bit‑exact activation splits at 1.25 GB/s with sub‑millisecond scheduling overhead. A graph‑based attribution method propagating credit across the full residual stream for safety‑relevant behaviors.

Activation StreamingThunderboltSafety ClassifierFeature Attribution
Abstract & PDF
PREPRINT WORKING PAPER

What Does MLX 4-bit Cost?

A controlled audit of 4‑bit MLX quantization cost for code generation on Apple Silicon. Measures perplexity, HumanEval/EvalPlus pass rates, and tokens/sec across quantization levels under a reproducible harness.

MLX4-bit QuantizationHumanEvalEvalPlus
Abstract & PDF
2011–2018 Archived

The pre‑lab years. More than twenty apps shipped since 2011, across domains with nothing in common — where the range came from.

EnergyTechBlockchainWearablesDevOps

More papers are in the works. Open a channel if you want to talk about any of them.