All Articles
18316 articles total
RippleBench: Capturing Ripple Effects Using Existing Knowledge Repositories
Introduction of RippleBench, a pipeline for analyzing the side-effects of model unlearning and editing in LLMs.
Towards Understanding What State Space Models Learn About Code
A systematic analysis of State Space Models (SSMs) in code understanding, introducing SSM-Interpret to analyze spectral shifts during fine-tuning.
Update on Ocean Observatories Initiative
An update on the status and operations of the Ocean Observatories Initiative.
It doesn't matter if it works
A discussion regarding the philosophical or technical implications of whether a system's internal workings matter if the end result is successful.
Reference-Driven Multi-Speaker Audio Scene Generation from In-the-Wild Priors
Introduces ScenA, a text-to-audio model that generates multi-speaker audio scenes with natural ambient noise and overlapping speech using flow-matching.
UBP2: Uncertainty-Balanced Preference Planning for Efficient Preference-based Reinforcement Learning
Presents UBP2, a model-based reinforcement learning approach that uses uncertainty-balanced preference planning to improve sample efficiency in preference-based RL.
Large-Scale OD Matrix Estimation with A Deep Learning Method
Proposes a hybrid deep learning and numerical optimization method for more accurate and real-time origin-destination (OD) matrix estimation in transport systems.
Recursive Joint Simulation in Games
Explores the use of recursive joint simulation between AI agents to achieve more cooperative outcomes in strategic game-theoretic settings.
Fully Geometric Multi-Hop Reasoning on Knowledge Graphs with Transitive Relations
Introduces GeometrE, a geometric embedding method for multi-hop reasoning on knowledge graphs that maps logical operations to pure geometric transformations.
PosterForest: Hierarchical Multi-Agent Collaboration for Scientific Poster Generation
Presents PosterForest, a training-free framework that uses a hierarchical multi-agent collaboration to automate scientific poster generation.
Structured Cognitive Loop for Behavioral Intelligence in Large Language Model Agents (Extended Revision: From Behavioral Architecture to Epistemic Accountability)
Proposes the Structured Cognitive Loop (SCL) architecture to improve accountability and success rates in LLM agents by separating cognition, memory, and control.
The Personalization Trap: How User Memory Alters Emotional Reasoning in LLMs
Analyzes how long-term user memory in LLMs can create the 'personalization trap,' where demographic profiles bias emotional reasoning and reinforce social inequalities.
The AirPods Effect
A discussion on the 'AirPods Effect', likely referring to how a specific product's success changes consumer behavior or industry standards.
Flip TABLE: storing arbitrary data in iNaturalist
A technical discussion about storing arbitrary data within the iNaturalist platform using a 'Flip TABLE' approach.
FDA advisors unanimously vote to approve Moderna's mRNA after agency drama
FDA advisors have unanimously voted to approve Moderna's mRNA vaccine following previous administrative delays.
As China looms, Taiwan makes more drones for defense and the US military
Taiwan is increasing drone production for national defense and for export to the US military amid rising tensions with China.
Anthropic's Claude Code Artifacts update brings live, shared dashboards and interactive workspaces to enterprises
Anthropic introduces 'Artifacts' for Claude Code, enabling the creation of live, shared, and interactive HTML workspaces for enterprise teams.
A Multi-Domain Benchmark for Detecting AI-Generated Text-Rich Images from GPT-Image-2
Researchers introduce a multi-domain benchmark to detect AI-generated text-rich images (like receipts and infographics) from GPT-Image-2.
Trade-offs in Medical LLM Adaptation: An Empirical Study in French QA
An empirical study on adapting LLMs for French medical QA, comparing continual pretraining (CPT) and supervised fine-tuning (SFT).
Correct Yourself, Keep My Trust: How Self-Correction and Social Connection Shape Credibility in Social Chatbots
Research on social chatbots suggests that self-correction is more effective at maintaining user trust than external corrections.