Independent AI systems architect building local first AI tools, governed inference infrastructure, and experimental research systems.
The interesting part is not prompting a model.
The interesting part is building the system around it: securely, observably, and in production.
Public Work
Vontra · local inference enablement
Helping make local inference more accessible through MLX model creation, testing, validation, and distribution.
TensorFold · Apple silicon / MLX local inference
Local-first runtime work for running sparse MoE language models on Apple silicon with MLX under tight memory budgets, with bounded resident memory, explicit paging, and runtime telemetry.



