Skip to content
View ashhart's full-sized avatar
🎯
Focusing
🎯
Focusing

Sponsoring

@MiaAI-Lab

Block or report ashhart

Block user

Prevent this user from interacting with your repositories and sending you notifications. Learn more about blocking users.

You must be logged in to block users.

Content in all repositories owned by your account will be closed.
Maximum 250 characters. Please don’t include any personal information such as legal names or email addresses. Markdown is supported. This note will only be visible to you.
Report abuse

Contact GitHub support about this user’s behavior. Learn more about reporting abuse.

Report abuse
ashhart/README.md

Ash Hart

I build and test local AI systems on Apple silicon and NVIDIA hardware. My work covers MLX, CUDA, mixed-hardware inference, local agents and runtime security.

The interesting part is not prompting a model.
The interesting part is building the system around it securely, observably and in production.

Build logs on X · Models on Hugging Face

Current experiments

Vontra  ·  local inference enablement
Helping make local inference more accessible through MLX model creation, testing, validation, and distribution.

TensorFold  ·  Apple silicon / MLX local inference
Local-first runtime work for running sparse MoE language models on Apple silicon with MLX under tight memory budgets, with bounded resident memory, explicit paging, and runtime telemetry.

Pinned Loading

  1. TensorFold TensorFold Public

    Run MoE LLMs on Apple Silicone via MLX that your Mac should not normally be able to run

    Python 40 4

  2. OpenBot OpenBot Public

    OpenBot - A reimplementation of GrokBot

    TypeScript 21 6

  3. Cortheon Cortheon Public

    A lightweight runtime that helps a small local model reason, discover, and complete work like a frontier model.

    Python 6 1

  4. ComputerUse ComputerUse Public

    Use computer use with tools like OMP or OpenCode.

    JavaScript 43 1