All Articles
17629 articles total
Has This Checkpoint Been Abliterated? A Two-Signal Audit and Its Failure Map
A new audit method combines activation gaps and weight-recovery energy to detect if an open-weight LLM checkpoint has had its refusal mechanisms removed (abliterated).
An Exploratory Study on LLM-Generated Code and Comments in Code Repositories
A study of LLM-generated code in repositories suggests that such code is frequently found in test cases and exhibits high levels of internal cloning, with little direct association with bugs.
SAB-LVLM: Significance-Aware Binarization for Large Vision-Language Models
SAB-LVLM introduces significance-aware binarization for Large Vision-Language Models to reduce memory and latency for resource-constrained deployment.
Rank-Then-Act: Reward-Free Control from Frame-Order Progress
Rank-Then-Act (RTA) is a framework for learning control policies from expert videos without environment rewards by using a VLM as a progress-based ordinal scorer.
Predicting Closed-Loop Performance of Latent World Models: Offline Checkpoint Selection for MPC and Model-Based RL Under Non-Markovian Rewards in LunarLander
Researchers introduce the Composite Reward Observability Fraction (CROF) to better select checkpoints for latent world models in Model-Based RL, significantly improving sample efficiency.
Full Bayesian Reinforcement Learning via LF-IBIS
The LF-IBIS algorithm enables full Bayesian Reinforcement Learning without requiring an explicit likelihood function by combining Approximate Bayesian Computation and Importance Sampling.
MedStreamBench: A Time-Aware Benchmark for Streaming and Proactive Medical Video Understanding
MedStreamBench is a new time-aware benchmark for medical video understanding, evaluating how models handle streaming data and proactive clinical alerting.
Decentralized Stochastic Subgradient-type Methods with Communication Compression for Nonsmooth Nonconvex Optimization
A new general framework for decentralized stochastic subgradient-type methods is proposed to handle nonsmooth nonconvex optimization with communication compression.
ProCal: Inference-Time Proposal Calibration for Open-Vocabulary Object Detection
ProCal is an inference-time calibration method that improves the localization quality of classification scores in open-vocabulary object detection.
AI Virtue: What is "Good" Knowledge in the Age of Artificial Intelligence?
An exploration of epistemic virtues in AI through digital humanities, questioning the nature of 'good' knowledge and generative value in the AI era.
Scene-Conditioned PINN-GNN for Multipath RF Maps: Cross-Scene Generation and In-Scene Completion
A unified RF map construction framework combining PINNs and GNNs is presented to improve multipath propagation modeling in wireless communication.
EPnG: Adaptive Expert Prune-and-Grow for Parameter-Efficient MoE Fine-tuning
EPnG is an adaptive prune-and-grow framework for MoE fine-tuning that reallocates LoRA capacity based on expert importance, achieving high performance with minimal parameters.
Lightweight Safe Reinforcement Learning for End-to-End UAV Navigation
A lightweight safe RL framework for UAV navigation is proposed, utilizing asymmetric convolutions and a Lagrangian-based PPO algorithm for collision-risk awareness.
Single-Channel EEG-Based Cognitive Load Assessment in Online Learning: A Hybrid Deep Learning Approach
A feasibility study on using single-channel EEG to assess cognitive load during online learning, including a reproducible evaluation pipeline and visualization tool.
Giant trees have no trouble pumping water to top branches
A discussion regarding the biological mechanism of how giant trees transport water to their highest branches.
Steam Controller Auto-Charge – pilot to magnetic charging puck using CV
A project pilot involving computer vision to facilitate magnetic charging for the Steam Controller.
Dispersion loss counteracts embedding condensation in small language models
A technical paper discussing how dispersion loss can mitigate the condensation effect in Small Language Models (SLMs).
Leanstral 1.5: Proof Abundance for All
A discussion about Leanstral 1.5, a tool or system related to formal verification and proof abundance.
Amsterdam invented the fire department
A historical piece regarding the origin of the Amsterdam fire department.
The circuit that lets your brain think and see
A discussion about a circuit design that enables brain-inspired thinking and visual perception.