Category
Persist AI agent context across sessions and tool switches
Fine-tune any LLM on a 4 GB GPU, one YAML config
LLM proxy with API translation and multi-backend routing
Local inference engine for DeepSeek V4 Flash/PRO and GLM 5.2
Agentic AI framework for the JVM
Open-source AI coding agent you can leave running locally.
Run huge LLMs on low-end GPUs with minimal VRAM
Accelerate AI applications with caching technology
Compress LLM context before it reaches the model
Self-host API execution for AI agents
Next-gen coding agent harness for efficient workflows
Smart AI Router with 3-Tier Fallback