LLM & AI infrastructure tools for API
LLM & AI infrastructure tools available for API, ranked by the AI Launch community. 7 products.
- 1
CUA-S1A System One Model for Computer Use
Cua gives AI agents computers they can use. It provides open-source desktop automation drivers, isolated cloud desktops (Cua Fleets), local macOS VMs (Lume), CUA-S1 specialist models for computer-use decisions, and Cua Bench for evaluating
π¬ 0Yβ² 95 on HNAI AgentsLLMs & InfrastructureFreemiumby AI Launch Team
- 2Jevstiller
Distill Jev into a local model β with a disagreement bound
Jevstiller distills a hosted Jev classifier into a local model, with a measured bound on how often the local model disagrees with Jev. It runs as a drop-in proxy, so batch jobs avoid a network call per answer while keeping answers close to
π¬ 0Yβ² 63 on HNLLMs & InfrastructureOpen Sourceby AI Launch Team
- 3
jevalsReplacing LLM judges with typed Jev decisions
jevals runs evals and guardrails for AI agents using Jev-style decision models instead of an LLM judge. All the checks for a trace, such as tool choice, groundedness, answer relevancy, prompt injection and PHI, go out as one request that co
π¬ 0Yβ² 47 on HNLLMs & InfrastructureAI AgentsOpen Sourceby AI Launch Team
- 4OpenLake
Storage engine for KV-cache offload and LLM training
OpenLake is a high-performance storage engine for LLM inference and GPU training. Its Infinity Core I/O Engine delivered 6.72 GiB/s writes and 11.55 GiB/s reads in the MLPerf Storage v3.0 Llama 3.1 8B checkpointing benchmark through its S3
π¬ 0Yβ² 35 on HNLLMs & InfrastructureOpen Sourceby AI Launch Team
- 5Lanes Link
A personal context MCP: one endpoint for every agent you use
Lanes Link is a private, self-hostable MCP server that connects your accounts, memory, skills and secrets to every AI agent you use, behind Access Profiles you manage. Connect services like Gmail, Calendar, Drive, Notion, Linear, Slack and
π¬ 0Yβ² 28 on HNLLMs & InfrastructureProductivityFreemiumby AI Launch Team
- 6
InstinctFlashHigh-Performance Serving Runtime for Robotics Models
InstinctFlash is a high-performance serving framework for robotics models from General Instinct (YC P26). It serves eight robotics model families, including pi05, LingBot-VLA, LingBot-VA and GR00T N1.7, through one runtime with Python and W
π¬ 0Yβ² 27 on HNLLMs & InfrastructureOpen Sourceby AI Launch Team
- 7OpenAPPA
Open-source deterministic guardrails that don't break agents
OpenAPPA is an information-flow policy engine that acts as a deterministic guardrail for LLM agents. It tracks how data flows instead of matching patterns, which the project says makes it fully resistant to data exfiltration from prompt inj
π¬ 0Yβ² 23 on HNAI AgentsLLMs & InfrastructureOpen Sourceby AI Launch Team
Related
Building one of these?
Launch it free