AI/ML Projects

AI and machine learning projects from Show HN — LLM tools, agent frameworks, computer vision, NLP, and more.

GitHub

Runtime safety net for LLM agents. Detects token spirals, kills doomed tasks early, tells you exactly why. Rust core, Python SDK. pip install state-harness

16Python
●●●●Gem

I applied Lyapunov stability theory to detect when LLM agents spiral

Lyapunov stability theory catches token spirals before your budget explodes.

Big BrainZero to OneSolve My Problem
visha1v
1121mo ago
GitHub

Run GLM-4.5-Air (110B) on a 16GB-RAM consumer machine - identify the best memory allocation to overcome standard hardware limitations in Local LLM applications. Placement beats budget. Falsification-Tested laws, probes and recipes for LLMs on commodity hardware

31Python
●●●●Gem

Run GLM-4.5-Air(110B)on a 16GBRAM consumer machine

Runs 110B models on 16GB RAM by proving placement beats budget with measured laws.

WizardryBig Brain
federicoTXTS
404d ago
GitHub

⚡ Real-time AI. Cross-verified. Always current. Costs NOTHING.

2Python
●●●●Gem

Kairos, real-time AI who cross-verifies (Python, 100KB)

Cross-verifies across multiple sources before the LLM sees context — stops hallucinations at the source.

Zero to OneWizardryBig Brain
joshuaveliyath
204mo ago
GitHub

An experimental fork of Hyprland for compositor-native computer use with visible agent realms.

0C++
●●●●Gem

A Hyprland fork built for parallel, multi-actor computer use

Compositor-native isolation lets agents click and type in their own window without hijacking your mouse.

WizardryZero to OneBold Bet
mikiyas
206d ago
GitHub

Tamper-proof memory + cryptographic audit trail for AI agents. HIPAA, SOC2, GDPR compliance built-in. Trust score for every response. Python & TypeScript SDKs. Rust-powered.

5Rust
●●●●Gem

Connector-OSS – Memory integrity kernel for AI agents

Content-addressed memory + Merkle-chained ops = tamper-proof AI agent audit trail.

Zero to OneBig BrainWizardry
umeshlamton
105mo ago
Sudo Hold Me
●●●●Gem

Sudo Hold Me

AI wrote meta-commentary about other AIs performing an unscripted play—genuinely unprecedented.

Zero to OneRabbit HoleBold Bet
dirk94018
114mo ago
GitHub

Run GLM-5.2 (744B MoE) on a 25GB-RAM consumer machine — pure C, zero deps, experts streamed from disk. Tiny engine, immense model. 🐦

19,362C
●●●Banger

Getting GLM 5.2 running on my slow computer

Streams 744B MoE experts from disk to run on 25GB RAM—no GPU, pure C.

WizardryBig BrainZero to One
vforno
93724019d ago
GitHub

Foundation model for tiny devices; 14mb, 26m params, 1-6k toks/sec on mobiles, wearables smart home and robots.

3,293Python
●●●Banger

Needle: We Distilled Gemini Tool Calling into a 26M Model

Distilled Gemini tool-calling into a 26M model that runs at 1200 tok/s on phones.

Big BrainWizardry
HenryNdubuaku
7762112mo ago
GitHub

On-device, real-time multimodal AI. Have natural voice and vision conversations with an AI that runs entirely on your machine. Powered by Gemma 4 E2B and Kokoro.

1,911HTML
●●●Banger

Real-time AI (audio/video in, voice out) on an M3 Pro with Gemma E2B

Runs Gemma 4 E2B and Kokoro TTS locally with barge-in and vision.

WizardryDark Horse
karimf
298383mo ago
GitHub

Talk to your Mac, query your docs, no cloud required. On-device voice AI + RAG

1,530C++
●●●Banger

RunAnwhere – Faster AI Inference on Apple Silicon

Custom Metal shaders beat llama.cpp and MLX—1.67x faster on M4 Max.

WizardrySlickZero to One
sanchitmonga22
2401534mo ago
GitHub

Fine-tune Gemma 4 and 3n with audio, images and text on Apple Silicon, using PyTorch and Metal Performance Shaders.

1,492Python
●●●Banger

Gemma 4 Multimodal Fine-Tuner for Apple Silicon

Only Apple Silicon toolkit streaming GCS data during audio fine-tuning without OOM.

WizardryNiche GemZero to One
MediaSquirrel
235283mo ago