Sign inLaunch

Open-source LLM & AI infrastructure tools

Open-source LLM & AI infrastructure tools launched on AI Launch, ranked by community upvotes. 8 products.

  1. 1
    Mini-AGI

    Dynamic continual learning model trained on 8GB VRAM

    mini-AGI is a byte-level language model that trains from scratch on a single 8 GB VRAM GPU and keeps learning from a continuous stream of text. Weights live on disk and are paged onto the GPU as needed, so model size is bounded by disk spac

    πŸ’¬ 0Yβ–² 277 on HNLLMs & InfrastructureResearch & DataOpen Sourceby AI Launch Team

  2. 2
    Lossless-memory

    A personal AI memory that never summarizes

    lossless-memory is a long-term memory system for a personal AI assistant that never summarizes. It keeps every raw line, timestamps everything, and searches by time first and words second, returning results in chronological order. A small i

    πŸ’¬ 0Yβ–² 67 on HNLLMs & InfrastructureProductivityOpen Sourceby AI Launch Team

  3. 3
    Jevstiller

    Distill Jev into a local model – with a disagreement bound

    Jevstiller distills a hosted Jev classifier into a local model, with a measured bound on how often the local model disagrees with Jev. It runs as a drop-in proxy, so batch jobs avoid a network call per answer while keeping answers close to

    πŸ’¬ 0Yβ–² 63 on HNLLMs & InfrastructureOpen Sourceby AI Launch Team

  4. 4
    jevals

    Replacing LLM judges with typed Jev decisions

    jevals runs evals and guardrails for AI agents using Jev-style decision models instead of an LLM judge. All the checks for a trace, such as tool choice, groundedness, answer relevancy, prompt injection and PHI, go out as one request that co

    πŸ’¬ 0Yβ–² 47 on HNLLMs & InfrastructureAI AgentsOpen Sourceby AI Launch Team

  5. 5
    OpenLake

    Storage engine for KV-cache offload and LLM training

    OpenLake is a high-performance storage engine for LLM inference and GPU training. Its Infinity Core I/O Engine delivered 6.72 GiB/s writes and 11.55 GiB/s reads in the MLPerf Storage v3.0 Llama 3.1 8B checkpointing benchmark through its S3

    πŸ’¬ 0Yβ–² 35 on HNLLMs & InfrastructureOpen Sourceby AI Launch Team

  6. 6
    AURA

    A Rust agent that investigates and fixes production incidents

    AURA is a production-tested SRE agent platform from Mezmo that investigates and helps fix production incidents. Specialist agents correlate traces, logs, metrics and deployment history to find root causes and recommend fixes, using the mode

    πŸ’¬ 0Yβ–² 28 on HNAI AgentsLLMs & InfrastructureOpen Sourceby AI Launch Team

  7. 7
    InstinctFlash

    High-Performance Serving Runtime for Robotics Models

    InstinctFlash is a high-performance serving framework for robotics models from General Instinct (YC P26). It serves eight robotics model families, including pi05, LingBot-VLA, LingBot-VA and GR00T N1.7, through one runtime with Python and W

    πŸ’¬ 0Yβ–² 27 on HNLLMs & InfrastructureOpen Sourceby AI Launch Team

  8. 8
    OpenAPPA

    Open-source deterministic guardrails that don't break agents

    OpenAPPA is an information-flow policy engine that acts as a deterministic guardrail for LLM agents. It tracks how data flows instead of matching patterns, which the project says makes it fully resistant to data exfiltration from prompt inj

    πŸ’¬ 0Yβ–² 23 on HNAI AgentsLLMs & InfrastructureOpen Sourceby AI Launch Team

Related

Building one of these?

Launch it free