FinRED

★ 80

Expert-guided red-teaming framework for financial LLM safety evaluation.

Modeloop

★ 80

Converts visual algorithms to microcontroller C code

A public repository providing a comprehensive index and standards for sign-language datasets.

STP-Lean

★ 80

A framework integrating Lean proof assistant for process-level reward signals in RL for theorem proving.

Large-scale training dataset with 238K episodes for embodied dialog navigation.

EPB

★ 80

A framework for interpreting Neural Combinatorial Optimization policies via distilled program portfolios.

ASyMOB

★ 80

Algebraic Symbolic Mathematical Operations Benchmark for evaluating LLM reasoning in math.

A frequency-domain framework for analyzing what State Space Models learn about code.

GeometrE

★ 80

A geometric embedding method for multi-hop reasoning on knowledge graphs with transitive relations.

ROLL

★ 80

An open-source RL platform used to implement the Spotlight system.

Rust implementation of the Mistral model for high-performance inference.

LiteLLM

★ 80

A proxy gateway that allows calling multiple LLM providers using a single OpenAI-compatible API.

Research proposing foundation models for Reinforcement Learning using synthetic MDPs.

A morphology-aware neural tokenizer and word embedder for the Turkish language.

scGTN

★ 80

Deep Siamese Graph Transformer Network for Single-cell RNA Sequencing Clustering

CURE

★ 80

Context management via Uncertainty-aware admission and Redundancy aware Eviction for TFMs

Unsloth

★ 80

A library for faster and more memory-efficient LLM fine-tuning.

Voronoi

★ 80

Implementation of structured adversarial camouflage using Voronoi diagrams for AI evasion.

An LLM-choreographed multi-agent world model for multi-view driving video generation.

Activation explainer for deception auditing in reasoning LLMs using hidden state decoding.