GPU-parallel robust control solver for nonlinear and NN dynamics with certified error bounds.

LOCOS

★ 80

Logit-Contribution Scoring for identifying non-literal retrieval heads in LLMs

Inkwell

★ 80

An RSS reader optimized for e-ink devices.

zkGolf

★ 80

Competitive optimization of formally verified circuits

Large-scale benchmark for LLVM compiler issue resolution containing 423 validated tasks.

A columnar log storage and query engine designed for high efficiency and scalability.

Manufact

★ 80

Cloud implementation of the Model Context Protocol (MCP) for AI agents.

SMT

★ 80

A black-box attack framework for jailbreaking function-calling LLMs via simulated moderation traces.

Agentic framework for iterative video retrieval and reasoning via soft query refinement.

Official code and videos for the Generalizable Data-efficient Agent (GenDa) framework.

A diagnostic egocentric video benchmark for evaluating embodied VLMs as runtime safety guards.

Self-GC

★ 80

A self-governing context management system for reducing token usage and maintaining agent state in long-horizon tasks.

An automated agentic pipeline for generating verifiable, deterministic reaction rules for chemistry synthesis.

MRPO

★ 80

Medical Reasoning-aware Policy Optimization for multimodal LLMs in clinical image reasoning.

HIG

★ 80

Framework for histogram-constrained image generation in diffusion models.

Uncertainty-guided synthetic context augmentation for semantic segmentation.

ZEBRA

★ 80

Zero-shot Entropy-Regularized Prompt Learning for Audio-Language Models.

A benchmark for measuring longitudinal psychometric stability (Mandate Salience Decay) in financial LLM agents.

Pglayers

★ 80

PostgreSQL extensions managed as stackable Docker layers.

Robust statistical AI-generated text detection framework.