All Dev Resources
3098 resources total
A curated collection of papers and resources regarding low-precision training for LLMs.
CSN
★ 85Code for Calibrating Siamese Network in Facial Action Unit Recognition
Orca-Bench
★ 85A benchmark for evaluating the capabilities of LLM agents in on-call operational scenarios.
BitBang
★ 85Reach machines behind NAT from a browser, no account required.
SymmGrid
★ 85A trajectory level augmentation framework to accelerate on-robot reinforcement learning.
ReCo
★ 85A reweighting method for Group Relative Policy Optimization (GRPO) to combat distributional concentration.
WhisperRec
★ 85Latent reasoning framework for efficient foundation recommendation models with high inference throughput.
AgentGUI
★ 85A locally hosted GUI for observing and steering long-running AI agents.
SimpleWikiSearch
★ 85A clean offline Wikipedia environment for reproducible agentic search evaluation.
AgenticCANN
★ 85Knowledge-augmented agentic evolution framework for automated Ascend C operator synthesis on NPUs.
TraceCoder
★ 85Code generation framework providing auditable and explainable repair histories for LLM agents.
ErsatzTV
★ 85A tool for creating a simulated cable TV experience from your own media library
Ollama
★ 85LLM inference layer for running open-source models locally.
LiveKit
★ 85An open-source real-time audio and video stack for building apps
OPI Abstraction
★ 85A standardization layer for Data Processing Units (DPUs) and Infrastructure Processing Units (IPUs).
Go LLM SDK
★ 85A Go SDK for streaming and tool-calling AI backends with a React frontend library.
RoboMME-Interference
★ 85A cross-session benchmark for measuring robot memory under interference.
PatchWorld
★ 85Gradient-free framework turning offline trajectories into executable Python world models.
CogTest
★ 85A principled benchmark designed to evaluate the cognitive habits of Large Reasoning Models (LRMs).
SheetasToken
★ 85A graph-enhanced framework for multi-sheet spreadsheet retrieval and understanding.