Hardware/Chips arXiv cs.AI

Efficient foundation decoders for fault-tolerant quantum computing

Introduces Neural Transfer Unification (NTU) and NTU-Transformer to efficiently train foundation decoders for fault-tolerant quantum computing across different code distances.

AI/ML arXiv cs.AI

Safe Autoregressive Image Generation with Iterative Self-Improving Codebooks

Introduces iterative self-improving codebooks to enhance the safety of autoregressive image generation by identifying and removing harmful mappings without human annotation.

Software Engineering Hacker News

Gossamer: a Rust-flavoured language with real goroutines and pause-free memory

Introduction of Gossamer, a new programming language inspired by Rust that features real goroutines and pause-free memory management.

Other Hacker News

The "Bizarre Headgear" Exhibit at the Sam Noble Museum Is Incredible

An exhibit of bizarre headgear at the Sam Noble Museum is highlighted as incredible.

Tech Business/VC TechCrunch

Novak Djokovic has a new job — advisor to private equity firm General Atlantic

Tennis star Novak Djokovic has become a global strategic advisor to the private equity firm General Atlantic.

Other Ars Technica

FCC accused of hiding Chairman Carr's messages with DOGE and Musk

The FCC is accused of withholding messages between Chairman Carr and Elon Musk's DOGE initiative.

Other Ars Technica

Netflix now requires every user profile to be tied to unique email address

Netflix implements a requirement for each user profile to be tied to a unique email address to prevent login sharing.

AI/ML arXiv cs.AI

On-board Remote-Sensing Foundation Models for Unsupervised Change Detection of Disaster Events

A new unsupervised change detection method for disaster events is proposed using Remote-Sensing Foundation Models (RSFMs) for onboard satellite processing.

Cybersecurity arXiv cs.AI

ShareLock: A Stealthy Multi-Tool Threshold Poisoning Attack Against MCP

ShareLock is introduced as a stealthy multi-tool threshold poisoning attack framework targeting the Model Context Protocol (MCP) for LLM agents.

AI/ML arXiv cs.AI

State Representation Matters in Deep Reinforcement Learning: Application to Energy Trading

Research demonstrates that state representation is critical for the success of Deep Reinforcement Learning agents in energy trading environments.

Software Engineering arXiv cs.AI

The Spec Growth Engine: Spec-Anchored, Code-Coupled, Drift-Enforced Architecture for AI-Assisted Software Development

The Spec Growth Engine is a framework for AI-assisted software development that uses a machine-readable spec graph to prevent context explosion and spec-code drift.

AI/ML arXiv cs.AI

NuclearQAv2: A Structured Benchmark for Evaluating Domain-Science Competence in Large Language Models

NuclearQAv2 is a new structured benchmark designed to evaluate the quantitative reasoning and domain-science competence of LLMs in nuclear engineering.

Other Hacker News

What Is a Nomogram and Why Would It Interest Me?

A discussion on Hacker News exploring the concept of nomograms and their utility in calculation and data visualization.

AI/ML arXiv cs.AI

Scaling Multi-Reference Image Generation with Dynamic Reward Optimization

Introduces OmniRef-Bench for evaluating multi-reference image generation and DyRef, a training framework to improve model performance in complex scenarios.

AI/ML arXiv cs.AI

XMSE-Aware Adaptive Empirical Bayes Estimation

Proposes an XMSE-aware mixed estimator that interpolates between maximum likelihood and Empirical Bayes shrinkage to improve estimation under kernel misspecification.

AI/ML arXiv cs.AI

In-Context Model Predictive Generation: Open-Vocabulary Motion Synthesis from Language Models to Physics

Presents In-Context Model Predictive Generation (ICMPG), a framework integrating LLM planning with physics simulation for realistic human motion synthesis.

AI/ML arXiv cs.AI

Auditing Framing-Sensitive Behavioral Instability in Large Language Models for Mental Health Interactions

Analyzes how different contextual framings impact the behavioral stability and internal representations of LLMs in mental health interaction scenarios.

AI/ML arXiv cs.AI

ReaORE: Reasoning-Guided Progressive Open Relation Extraction Empowered by Large Reasoning Models

Introduces ReaORE, a reasoning-guided progressive framework that uses coarse-to-fine reasoning to improve Open Relation Extraction.

AI/ML arXiv cs.AI

Where Do Models Find Happiness? Emotion Vectors in Open-Source LLMs

Investigates emotion vectors in open-weight LLMs like Apertus and Gemma, discovering how valence representations emerge differently across model depth.

AI/ML arXiv cs.AI

Decision-Aligned Evaluation of Uncertainty Quantification

Proposes a decision-aligned evaluation framework and prior-weighted utility metrics to better assess uncertainty quantification in machine learning.