Skip to content
View ashhart's full-sized avatar
🎯
Focusing
🎯
Focusing

Block or report ashhart

Block user

Prevent this user from interacting with your repositories and sending you notifications. Learn more about blocking users.

You must be logged in to block users.

Maximum 250 characters. Please don’t include any personal information such as legal names or email addresses. Markdown is supported. This note will only be visible to you.
Report abuse

Contact GitHub support about this user’s behavior. Learn more about reporting abuse.

Report abuse
ashhart/README.md

Ash Hart

Independent AI systems architect building local first AI tools, governed inference infrastructure, and experimental research systems.

The interesting part is not prompting a model.
The interesting part is building the system around it: securely, observably, and in production.


Toolbox

AI Agents Local LLMs RAG MCP Model Routing Agent Safety Operational Intelligence Stateful Systems Rust Python TypeScript


Public Work

Vontra  ·  local inference enablement
Helping make local inference more accessible through MLX model creation, testing, validation, and distribution.

TensorFold  ·  Apple silicon / MLX local inference
Local-first runtime work for running sparse MoE language models on Apple silicon with MLX under tight memory budgets, with bounded resident memory, explicit paging, and runtime telemetry.

Pinned Loading

  1. TensorFold TensorFold Public

    Run MoE LLMs on Apple Silicone via MLX that your Mac should not normally be able to run

    Python 38 4