Pglayers

★ 80

PostgreSQL extensions managed as stackable Docker layers.

Robust statistical AI-generated text detection framework.

BGT-ADA

★ 80

Budget-adaptive routing implementation for optimizing edge-cloud inference latency and accuracy.

Code-mixed benchmark for evaluating LLMs on Romanized Indic-English instructions.

Frond

★ 80

A frontend runtime for your app's dependency graph.

ELEVATE

★ 80

A framework for efficient GenAI-driven avatar tutors with local-first execution on consumer-grade hardware.

Cursor

★ 80

An AI-powered code editor used by a significant portion of Rivian's software team.

AxDafny

★ 80

A verifier-guided repair framework for generating verified code in Dafny.

MARS

★ 80

Modality-Agnostic Refusal Steering for training-free multimodal safety in LLMs.

Evo-PI

★ 80

A framework for improving medical visual question answering through evolving principle-guided supervision.

SAGE

★ 80

Self-correcting Autonomous Grounded Experimenter for reliable scientific research agents.

An LLM-based multi-agent workflow for refactoring software into HLS-compatible programs.

LabGuard

★ 80

A language-to-execution safety suite for grounding lab rules into runtime guards for AI agents.

A suite of simulation environments probing LLM belief trajectories under multi-turn evidence accumulation.

Benchmarks (AIOps2025, RCA100) for evaluating LLM agents in microservice failure diagnosis.

A framework for semantic-aware, physics-informed, and geometry-grounded weather video synthesis.

High-performance inference engine implementing HARD-KV for efficient KV cache compression.

BREIT

★ 80

Modular framework for 3D MF-EIT stroke reconstruction including forward solvers and D-bar implementation.

A five-stage PDP/PEP for agent tool calls to provide deterministic value authorization gating.

A benchmark for evaluating temporal visual reasoning in video-to-code generation.