Software Engineering Hacker News

The 80% Problem: The Last 20% Is Where the Engineer Used to Live

A community discussion on Hacker News exploring the shift in engineering where the 'final 20%' of a project's complexity is increasingly handled by automated tools rather than manual engineering effort.

Software Engineering Hacker News

Alternatives to Nested If Function

A technical discussion regarding alternatives to using deeply nested IF functions in spreadsheets or code to improve readability and maintainability.

Other TechCrunch

Watch out, Amazon: the Kobo eReader now has a Goodreads rival

Kobo eReaders now integrate with StoryGraph, providing users an alternative to Amazon's Goodreads for tracking reading progress and statistics.

Hardware/Chips The Verge

Sony’s next-gen PlayStation will go ‘beyond the living room’

Sony suggests that the next-generation PlayStation will move beyond the living room, hinting at a new handheld or more portable gaming experiences.

Hardware/Chips The Verge

OpenAI is teasing new hardware… for Codex

OpenAI is collaborating with Work Louder to release a hardware device featuring shortcuts for Codex to streamline AI-powered coding workflows.

Tech Business/VC Ars Technica

Google warns EU's plans to weaken its monopoly could expose user data

Google argues that EU mandates to share search data with competitors to curb its monopoly could compromise user privacy.

Hardware/Chips Ars Technica

Quantum computing startup says it will leapfrog everybody

A quantum computing startup claims it will surpass competitors, though the goal requires a significant leap in existing hardware capabilities.

Tech Business/VC Ars Technica

Kalshi sues Illinois over new tax on prediction market sports bets

Prediction market platform Kalshi is suing the state of Illinois over a new tax imposed on sports-based prediction market bets.

AI/ML arXiv cs.AI

AgentPSO: Evolving Agent Reasoning Skill via Multi-agent Particle Swarm Optimization

Researchers introduce AgentPSO, a framework that uses particle swarm optimization to evolve reasoning skills in multi-agent LLM systems without updating model parameters.

AI/ML arXiv cs.AI

Distilling Answer-Set Programming Rules from LLMs for Neurosymbolic Visual Question Answering

A new approach for neurosymbolic Visual Question Answering (VQA) that distills Answer-Set Programming (ASP) rules from LLMs to improve interpretability and adaptability.

Tech Business/VC TechCrunch

Waymo and Uber quietly part ways in Phoenix

Waymo and Uber have ended their three-year partnership in Phoenix, Arizona.

AI/ML arXiv cs.AI

Chronic Kidney Disease Prognosis Prediction Using Transformer

Researchers introduced ProQ-BERT, a transformer-based framework for predicting Chronic Kidney Disease progression using multi-modal electronic health records.

AI/ML arXiv cs.AI

Towards Benign Memory Forgetting for Selective Multimodal Large Language Model Unlearning

The SMFA framework and S-MLLMUn Bench are proposed to enable selective unlearning of privacy-sensitive information in multimodal LLMs without degrading general capabilities.

AI/ML arXiv cs.AI

Health-ORSC-Bench: A Benchmark for Measuring Over-Refusal and Safety Completion in Health Context

Health-ORSC-Bench is a new large-scale benchmark to measure over-refusal and safe completion quality of LLMs in healthcare contexts.

AI/ML arXiv cs.AI

GAIA: A Data Flywheel System for Training GUI Test-Time Scaling Critic Models

The GAIA system introduces a data flywheel to train intuitive critic models that improve the test-time scaling performance of GUI agents.

Cybersecurity arXiv cs.AI

Just Ask: Curious Code Agents Reveal System Prompts in Frontier LLMs

The JustAsk framework demonstrates that autonomous code agents can be used to systematically recover hidden system prompts from frontier LLMs.

AI/ML arXiv cs.AI

CausalFlip: A Benchmark for LLM Causal Judgment Beyond Semantic Matching

CausalFlip is a new benchmark designed to test whether LLMs are performing true causal reasoning or merely relying on semantic matching.

AI/ML arXiv cs.AI

Conservative Equilibrium Discovery in Offline Game-Theoretic Multiagent Reinforcement Learning

COffeE-PSRO is introduced as a conservative equilibrium discovery method for offline game-theoretic multiagent reinforcement learning.

AI/ML arXiv cs.AI

SEA-TS: Self-Evolving Agent for Autonomous Code Generation of Time Series Forecasting Algorithms

SEATS is a self-evolving agent framework that autonomously generates and optimizes code for time series forecasting algorithms.

AI/ML arXiv cs.AI

Algorithms for Deciding the Safety of States in Fully Observable Non-deterministic Problems: Technical Report

A new policy-iteration algorithm, iPI, is presented to decide the safety of states in non-deterministic problems with guaranteed polynomial worst-case runtime.