Open-source LLM & AI infrastructure tools
Open-source LLM & AI infrastructure tools launched on AI Launch, ranked by community upvotes. 8 products.
- 1
Mini-AGIDynamic continual learning model trained on 8GB VRAM
mini-AGI is a byte-level language model that trains from scratch on a single 8 GB VRAM GPU and keeps learning from a continuous stream of text. Weights live on disk and are paged onto the GPU as needed, so model size is bounded by disk spac
π¬ 0Yβ² 277 on HNLLMs & InfrastructureResearch & DataOpen Sourceby AI Launch Team
- 2
Lossless-memoryA personal AI memory that never summarizes
lossless-memory is a long-term memory system for a personal AI assistant that never summarizes. It keeps every raw line, timestamps everything, and searches by time first and words second, returning results in chronological order. A small i
π¬ 0Yβ² 67 on HNLLMs & InfrastructureProductivityOpen Sourceby AI Launch Team
- 3Jevstiller
Distill Jev into a local model β with a disagreement bound
Jevstiller distills a hosted Jev classifier into a local model, with a measured bound on how often the local model disagrees with Jev. It runs as a drop-in proxy, so batch jobs avoid a network call per answer while keeping answers close to
π¬ 0Yβ² 63 on HNLLMs & InfrastructureOpen Sourceby AI Launch Team
- 4
jevalsReplacing LLM judges with typed Jev decisions
jevals runs evals and guardrails for AI agents using Jev-style decision models instead of an LLM judge. All the checks for a trace, such as tool choice, groundedness, answer relevancy, prompt injection and PHI, go out as one request that co
π¬ 0Yβ² 47 on HNLLMs & InfrastructureAI AgentsOpen Sourceby AI Launch Team
- 5OpenLake
Storage engine for KV-cache offload and LLM training
OpenLake is a high-performance storage engine for LLM inference and GPU training. Its Infinity Core I/O Engine delivered 6.72 GiB/s writes and 11.55 GiB/s reads in the MLPerf Storage v3.0 Llama 3.1 8B checkpointing benchmark through its S3
π¬ 0Yβ² 35 on HNLLMs & InfrastructureOpen Sourceby AI Launch Team
- 6
AURAA Rust agent that investigates and fixes production incidents
AURA is a production-tested SRE agent platform from Mezmo that investigates and helps fix production incidents. Specialist agents correlate traces, logs, metrics and deployment history to find root causes and recommend fixes, using the mode
π¬ 0Yβ² 28 on HNAI AgentsLLMs & InfrastructureOpen Sourceby AI Launch Team
- 7
InstinctFlashHigh-Performance Serving Runtime for Robotics Models
InstinctFlash is a high-performance serving framework for robotics models from General Instinct (YC P26). It serves eight robotics model families, including pi05, LingBot-VLA, LingBot-VA and GR00T N1.7, through one runtime with Python and W
π¬ 0Yβ² 27 on HNLLMs & InfrastructureOpen Sourceby AI Launch Team
- 8OpenAPPA
Open-source deterministic guardrails that don't break agents
OpenAPPA is an information-flow policy engine that acts as a deterministic guardrail for LLM agents. It tracks how data flows instead of matching patterns, which the project says makes it fully resistant to data exfiltration from prompt inj
π¬ 0Yβ² 23 on HNAI AgentsLLMs & InfrastructureOpen Sourceby AI Launch Team
Related
Building one of these?
Launch it free