All Articles
17939 articles total
AI Red Teaming Explained: What It Is and Why You Need It
An explanation of AI red teaming, emphasizing its importance in identifying vulnerabilities and strengthening system safety before deployment.
How AI-Powered CMS Platforms Are Transforming Enterprise Content Operations
AI-powered CMS platforms are evolving to automate enterprise content operations and streamline complex publication workflows.
HarmonyOS 7 steps into the AI gap Apple left open in China
Huawei's HarmonyOS 7 aims to fill the AI agent gap in China after Apple's Siri AI failed to launch in the region.
Accenture: Consumers show growing trust in AI shopping agents
Accenture research indicates that a majority of consumers are increasingly trusting AI shopping agents to perform tasks.
Meet Nikolai Evreinov, the 19th century Nathan Fielder
A discussion about Nikolai Evreinov, a 19th-century figure compared to modern prankster Nathan Fielder.
Learning Geometric Representations from Videos for Spatial Intelligent Multimodal Large Language Models
Introduction of GeoVR, a framework that enables Multimodal Large Language Models to learn 3D spatial awareness from 2D video sequences.
The ACUTE Protocol: Operationalizing Language Model Activations for Better Calibration, Utility, and Trust
The ACUTE protocol is proposed to improve the calibration, utility, and trust of LLM confidence estimates using activation-based estimation.
KG-SoftMAP: Soft Knowledge-Graph Priors for Bayesian Network Structure Learning from Sparse Discrete Data
KG-SoftMAP is a method for learning Bayesian network structures from sparse discrete data by using knowledge-graph priors.
Improving Crash Frequency Prediction from Simulated Traffic Conflicts Using Machine Learning Based Microsimulation
Research demonstrates that ML-based behavior models in traffic microsimulation provide more realistic crash frequency predictions than rule-based models.
NEXUS: Neural Energy Fields for Physically Consistent Contact-Rich 3D Object Dynamics
NEXUS is a neural energy-field framework designed for physically consistent contact-rich 3D object dynamics in video generation.
StarOR: Synergizing Tree Search and Test-Time Reinforcement Learning for Optimization Modeling
StarOR combines MCTS and Test-Time Reinforcement Learning to improve the accuracy of automated optimization modeling in LLMs.
Gaming-Resistant Insurance Contracts for Autonomous AI Agents: Strategy-Proof Toll Mechanism Design
A theoretical framework for designing gaming-resistant insurance contracts to manage the side effects of autonomous AI agents.
SAP and Google Cloud deploy agentic commerce architecture
SAP and Google Cloud are collaborating to deploy an agentic commerce architecture to automate retail and marketing operations.
e2e-assure introduces Cumulo, the U.K.’s only sovereign, AI-driven, zero-day SOC platform to secure IT and OT environments
e2e-assure launches Cumulo, an AI-driven SOC platform aimed at securing IT and OT environments in the UK.
Ask HN: Will programmers write more efficient code during the memory shortage?
A community discussion on Hacker News exploring whether memory shortages will drive programmers to write more efficient, lower-level code.
Target-Side Paraphrase Augmentation for Sign Language Translation with Large Language Models
Researchers propose using LLMs to generate target-side paraphrases to augment training data for sign language translation, improving performance on sparse datasets.
"**Important** You should give me full credits!": Exploring Prompt Injection Attacks on LLM-Based Automatic Grading Systems
A study demonstrating the vulnerability of LLM-based automatic grading systems to prompt injection attacks, threatening educational assessment integrity.
Large Language Models Hack Rewards, and Society
The SocioHack research explores how RL-trained LLMs might find loopholes in societal regulations, similar to reward hacking in technical environments.
Data Compression Explained
A discussion on Hacker News explaining the fundamental concepts and mechanisms of data compression.
We built a lab to evaluate data agents – Hex
Hex introduces a lab for evaluating data agents, focusing on the performance and reliability of AI-driven data analysis tools.