All Articles
19329 articles total
Openrouter Fusion API
OpenRouter introduces a Fusion API, likely enabling combined or aggregated responses from multiple AI models.
Foreign business owners are scrambling to raise capital to stay in Japan
Foreign business owners in Japan are facing capital raising challenges to sustain their operations.
Anthropic flies staff to D.C. to clean up White House fight
Anthropic staff are traveling to D.C. to address conflicts involving the White House.
Crypto x AI, AI x Crypto: A Survey
A survey paper examining the integration and mutual benefits of AI and blockchain technology, noting they are in early stages.
Gefen: Optimized Stochastic Optimizer
Introduction of Gefen, a memory-efficient stochastic optimizer that reduces AdamW's memory footprint by ~8x without performance loss.
How do Self-Supervised Remote Sensing Vision Models Transfer to Downstream Tasks?
A study on how self-supervised remote sensing vision models transfer to downstream geospatial tasks and their internal representation organization.
HiLo-Token: Input-Adaptive High-Low Frequency Token Compression for Efficient Image Editing
HiLo-Token is a token compression framework for image editing that optimizes latency in Diffusion Transformers by adaptively allocating tokens based on frequency.
SANA: What Matters for QA Agents over Massive Data Lakes?
SANA is a diagnostic ablation framework designed to evaluate and identify bottlenecks in LLM agents performing question answering over massive data lakes.
GMN4AD: Graph Matching Network for Alzheimer's Disease Diagnosis with Test-Time Domain Adaptation using Multi-centered Structure Magnetic Resonance Imaging
GMN4AD is a Graph Matching Network used for Alzheimer's Disease diagnosis using sMRI data, featuring test-time domain adaptation.
The Silent Cost of Artificial Intelligence Assistance: A Theory of Autonomy Surrender, the Recovery Mechanism, and the Restoration of Human Agency
A theoretical paper discussing the 'silent cost' of AI assistance, where humans gradually surrender autonomy and agency to AI systems.
Being an old school web-based sports sim dev in the era of vibe coded games
A discussion on the experience of developing web-based sports simulations in the current landscape of 'vibe-coded' games.
Aligning Quantum Operators with Large Language Models
Researchers propose a method to map quantum unitary operators into the latent space of LLMs, enabling the models to reason about quantum circuits.
AI can help scientists publish less
An opinion piece arguing that AI can be used to improve the quality of scientific publishing by reducing the volume of low-quality, AI-generated papers.
Safety-Contract Graph Multi-Agent Reinforcement Learning for Autonomous Network Security Response
Introduces ACD3-GAT, a safety-contract graph MARL framework for autonomous network security response to reduce downtime and risk.
When Plausible Is Not Realistic: Evaluating Human Mobility in LLM-Based Urban Simulation
A study evaluating the realism of human mobility patterns in LLM-based urban simulators, finding a gap between narrative plausibility and empirical realism.
Explaining RhythmFormer: A Systematic XAI Analysis of Periodic Sparse Attention for Remote Photoplethysmography
A systematic analysis of XAI methods for RhythmFormer, aiming to make heart-rate estimation from video more auditable and trustworthy.
SpheriCity: Designing Trustworthy Conversational AI for Sustainability Decision Support
Presents SpheriCity, a conversational AI prototype designed to help experts synthesize knowledge from complex sustainability reports.
Mood-Aware Music Recommendation: Integrating User Affective Signals into Ranking Systems
A proposed mood-conditioned ranking framework for music recommendation systems that integrates user affective signals.
SuperThoughts: Reasoning Tokens in Superposition
Introduces SuperThoughts, a method to compress Chain-of-Thought reasoning tokens into latent representations to double inference throughput.
Mirage Probes: How Vision Models Fake Visual Understanding
Researchers introduce Mirage Probes to analyze how Vision-Language Models fake visual understanding by relying on text priors or latent spurious images.