中文

Agent infrastructure

259 projects · page 3 / 3

201-259 of 259
201
webai2api
Concurrency gateway converting web-based AIs such as LMArena and Gemini into OpenAI-compatible APIs via Camoufox anti-detection browsing with multi-window concurrency and account isolation.
202
mxc
Microsoft cross-platform sandboxed execution container provides policy-driven layered isolation for untrusted code, offering multiple backends and containment controls for security-conscious developers and platforms.
203
mindwalk
Visualization tool replaying coding-agent sessions on a 3D codebase map by locally parsing logs to show search, read, and write trajectories.
204
kanban
Local kanban web app that creates isolated worktrees for each task to run CLI agents in parallel, managing dependency chains and supporting diff review.
205
waza
Supplies a CLI and framework for creating, testing, and measuring Agent Skills, enabling skill authors to benchmark quality and compare effectiveness across different models.
206
NeuroAPI
Open one-click installers that connect Codex CLI and Claude Code to the NeuroAPI multi-model gateway.
207
skillsgate
Desktop application for browsing, installing, and managing AI agent skills, supporting more than twenty agents with local-first storage and skill synchronization.
208
flashkda
High-performance CUDA kernel library implementing Kimi Delta Attention to accelerate linear-attention computation during large-language-model inference for researchers and engineers optimizing throughput.
209
agent-as-a-router
The official implementation of Agent-as-a-Router for coding tasks, providing benchmark data, reference outputs, and gateway-level integration demos for backend model routing.
210
gym
A library providing environments and infrastructure for evaluating and improving models and agents at scale. Supports reproducible benchmarks and reinforcement-training integrations across multiple frameworks for AI researchers.
211
agent-memory-leaderboard
Public evaluation platform for agent long-term memory, defining a unified protocol with multi-track benchmarks and capability profiles to enable comparable leaderboard rankings.
212
gomodel
Offers a Go-based AI gateway unifying OpenAI- and Anthropic-compatible APIs across many providers, with intelligent routing, failover, streaming, observability, guardrails, and cost tracking.
213
halo
Desktop application for optimizing AI agents using production traces and reinforcement-style analysis. Imports browsing replays, ranks failure modes, and generates actionable fix suggestions for engineers improving hierarchical agent loops.
214
best-claude-hud
High-performance Rust-powered statusline HUD for Claude Code that displays current model, git branch, and usage information directly inside the developer's terminal workflow.
215
beellama.cpp
llama.cpp fork focused on local LLM inference efficiency, using KVarN, KV-cache precision tailoring, and low-bit quantization to support longer context with better precision in the same VRAM.
216
exploitgym
A large-scale realistic benchmark built from real-world vulnerabilities, including user-space and kernel cases, for evaluating AI agents' ability to develop working exploits.
217
numbat
An endpoint observability platform for security teams monitoring AI agent activity, combining on-device detection, optional pre-action blocking, and forensic session reconstruction to investigate automation behavior and policy violations.
218
tty7
A pure-Rust terminal workbench built with Zed gpui and Alacritty VT core, offering shells, persistent sessions, SSH access, and coding-agent collaboration for developers.
219
skill-up
An evaluation harness for Agent Skills that runs declarative cases across engines, judges results, and drives iterative skill improvement.
220
oc-go-cc
A Go proxy gateway that lets Claude Code run through OpenCode Go, Zen, Bedrock and OpenRouter with routing and failover.
221
one-api-pro
An enterprise AI API gateway unifying 30+ model providers behind OpenAI-compatible endpoints, with token controls, billing, failover and active-active clustering.
222
smolvm
Supplies open-source sandbox infrastructure for AI workloads with a unified API over Firecracker, QEMU, and libkrun, supporting persistent microVMs, browser access, and file sharing.
223
pideck
An Electron-based desktop workspace for developers managing multiple pi Agent sessions across local project directories, with unified history browsing, session restore, Git integration, built-in terminal, model settings, and plugin support.
224
ds4-control
macOS menu-bar application for managing local ds4 model services with fast DeepSeek V4 Pro and Flash variants. It handles model downloads, service monitoring, and launching coding assistants with million-token context support.
225
kero
A native macOS terminal workspace combining tabs, splits, a browser, and a file tree, with background AI agents and unified state management for collaborative command-line work.
226
hermes-hud
Offers a terminal-based consciousness monitor for persistent-memory Hermes agents, visualizing memory, skills, projects, health status, and error-correction activity in real time for operators.
227
WindowsAgentArena
Creates a scalable benchmarking platform for multimodal Windows agents, providing authentic operating-system environments and large-scale parallel evaluation for researchers testing desktop automation.
228
opencode-manager
Mobile-first web interface for managing multiple OpenCode AI agents from any device, offering session chat, file management, Git integration and task scheduling in a deployable PWA.
229
deltafin
A single-file inference program for running the full Kimi K3 model on consumer hardware, plus an OpenAI-compatible local API server for chat and coding agents.
230
re_gent
Local version-control and auditing tool for AI coding agents that records agent actions, attributes code changes to specific prompts, and lets developers inspect individual steps with surrounding context for debugging.
231
chatgpt2api
Implements an OpenAI-compatible gateway by reverse-engineering ChatGPT web protocols, supporting account pool management, text and GPT-Image models, plus batch image generation and editing.
232
pandaprobe
An open-source agent engineering platform with tracing, evaluations, metrics, and dashboards to debug and improve AI agents.
233
agentacct
Local-first dashboard for coding agents such as Claude Code and Codex, breaking each task into work receipts with tools used, files changed, tests run, time, and token costs.
234
llama.cpp
Maintains a fork of the llama.cpp local LLM inference engine, supporting low-bit model formats and multiple hardware backends for developers building local inference services.
235
opencodex
Local gateway and dashboard for Codex Desktop that handles account-pool routing, authentication, and task visualization for teams operating multiple Codex accounts.
236
nadirclaw
An open-source LLM router and cost optimizer that automatically sends simple prompts to cheap or local models and complex ones to premium models via an OpenAI-compatible proxy.
237
kooky
Minimal modern terminal designed for AI coding workflows, offering sidebar workspaces, horizontal and vertical split panes, one-click agent launching, and live per-agent activity and workspace state.
238
codex-relay
Relay service with a mobile companion letting users remotely follow and control local Codex sessions from a phone while computation remains on the computer.
239
anolisa
An operating layer for AI agent workloads combining an AI-native terminal, token-saving compression, eBPF observability, memory and sandboxed secure execution.
240
agent-inspect
Helps TypeScript and Node.js developers inspect AI agent runs locally by turning manual steps, tool calls, model calls, logs, failures, and durations into readable terminal execution trees.
241
agterm
Native terminal built for parallel multi-agent workflows, organizing work through workspaces and sessions while exposing a full control API for dispatching coding tasks efficiently.
242
emilia-protocol
Protocol and gateway acting as a consequence firewall for machine actions, verifying exact authority before money, code, permission, or infrastructure changes and producing independently verifiable receipts.
243
robodojo
An evaluation benchmark for general-purpose robot manipulation policies, combining simulated tasks with real-world validation procedures to help robotics researchers compare manipulation strategies reproducibly.
244
cc-lens
Local analytics dashboard for Claude Code users that visualizes session costs, model usage, project trends, and task-level details. Runs entirely offline without cloud services or telemetry for privacy-preserving usage analysis.
245
locate-anything.cpp
C++ ggml port of Nvidia's LocateAnything-3B open-vocabulary detection model enabling fast object localization on CPU and GPU for developers building vision applications.
246
bigmoeonedge
An inference engine built on stock llama.cpp that runs oversized MoE models on memory-constrained phones by streaming only routed experts from flash without quality loss.
247
cl-bench
A benchmark series for context learning that evaluates reasoning and learning ability through workplace and everyday tasks with thousands of detailed annotation rules.
248
opencode-with-claude
OpenCode plugin that routes calls through a Claude Max or Pro subscription via Meridian, automatically managing local proxy lifecycles for conflict-free multi-instance use.
249
grokcli-2api
A gateway service that converts authenticated Grok logins into endpoints compatible with multiple LLM API specifications, offering multi-worker processing and a web admin panel for operators.
250
vllm-moet
vLLM patch with hand-written SM120 SASS kernels combining 2-bit MoE experts and an FP4 delta cache to recover precision, enabling frontier MoE models to run on consumer Blackwell GPUs.
251
yepanywhere
Self-hosted web interface for remotely monitoring and controlling existing Claude and Codex CLI sessions from phones or other devices. Supports push notifications, tool approvals, file uploads, and device management without accounts or databases.
252
amd-strix-halo-vllm-toolboxes
Containerized LLM inference toolboxes for developers on AMD Strix Halo hardware, tracking versioned releases and supporting both single-machine deployment and multi-node inference clusters.
253
vibearound
Unified launch and management hub for multiple AI coding agents such as Claude Code, Codex CLI, and Gemini CLI, supporting parallel sessions across desktop, mobile, and messaging.
254
obelisk
Indexes past agent sessions, subagents, and workflows into a unified local memory store that agents can query and users can browse for continuity.
255
krasis
A hybrid LLM inference runtime for running very large models on consumer GPUs with limited VRAM. Uses GPU prefill, decoding, and hot-cold expert management to enable local execution for individual practitioners.
256
skills-npm
A command-line installer that fetches agent skills from npm packages and links them into multiple agent configuration directories, helping developers keep skill definitions synchronized across tools.
257
harness-score
Offline evaluation toolkit that measures the maturity of coding-agent harnesses in seconds, scoring reliability practices and returning a prioritized remediation checklist for developers.
258
avibe
Local-first agent operating system that drives official Claude Code, Codex and OpenCode instances from browsers or chat apps while unifying workspace management on the user machine.
259
tabbit-toy
Local gateway for users converting Tabbit access into OpenAI-compatible APIs, bundling membership authentication and a browser extension for one-click cookie extraction to call Claude and GPT models.