Opus 5.5 agents discover two room-temperature magnetic semiconductor candidates
Article URL: https://www.vals.ai/blogs/room-temperature-magnetic-semiconductors Comments URL: https://news.ycombinator.com/item?id=49970667 Points: 132 # Comments: 109
A compact digest of model releases, papers, agentic systems, coding automation, infra, security, and high-signal developer discourse. Feed candidates are ranked locally; the scheduled Hermes worker can add qualitative synthesis.
Article URL: https://www.vals.ai/blogs/room-temperature-magnetic-semiconductors Comments URL: https://news.ycombinator.com/item?id=49970667 Points: 132 # Comments: 109
Article URL: https://reflection.ai/blog/introducing-beam Comments URL: https://news.ycombinator.com/item?id=49969183 Points: 234 # Comments: 63
Article URL: https://www.techspot.com/news/114091-florida-woman-used-claude-diary-anthropic-reported-shoot.html Comments URL: https://news.ycombinator.com/item?id=49961057 Points: 450 # Comments: 378
Hey HN, Today, we're launching selfbench.dev, an open-source tool that lets you create and run evals automatically from your PRs. Every benchmark with sufficient trust eventually gets benchmaxxed (Goodhart's law) - the labs are incentivized to maximize their scores on that benchmark, which isn't predictive on whether it'll actually work within your setup.…
On Friday I gave the closing keynote at the WeAreDevelopers World Congress North America in San Jose. I tied together the key trends from the past year into a chronological exploration of everything that happened in 2026. The video is on YouTube ; here are my annotated slides and notes to accompany the talk. And as an annotated presentation : # I'm going…
I do a lot of work using claude and codex. I use plenty of sub-agents. But nothing balances bang for the buck as good as deepseek v4.1 flash for me. So I thought, why not use the deepseek harness as MCP and drive it from claude as an orchestrator. I am probably not alone with this idea, but I haven't found something comparable. I wanted to fan out sub-age…
Article URL: https://arxiv.org/abs/2602.11243 Comments URL: https://news.ycombinator.com/item?id=49968612 Points: 1 # Comments: 0
I have been working on Rashomon, an open-source execution recorder for AI coding agents. The basic idea is that the agent transcript is not ground truth. Rashomon keeps its own record of what happened and compares it against the agent’s account. For Claude Code, it records things like: * shell commands, exit codes, and whether they may have written files…
Hi guys, Posted this ( https://news.ycombinator.com/item?id=48832797 ) on HN a few months ago when it was a different product (and closed source). It was originally a way of managing multiple agents in parallel (like TMUX, but with a nicer chat UI). But I found that I was only able to work with 3 or 4 agents simultaneously without information overload, so…
Article URL: https://threadnote.io/whats-new/articles/graphmem-agent-continuation-study/ Comments URL: https://news.ycombinator.com/item?id=49970034 Points: 2 # Comments: 2
Article URL: https://github.com/cleuton/MetaAgent Comments URL: https://news.ycombinator.com/item?id=49967468 Points: 1 # Comments: 0
Article URL: https://github.com/Agilno-Tech/rivet/ Comments URL: https://news.ycombinator.com/item?id=49965688 Points: 1 # Comments: 1
Article URL: https://github.com/Nero7991/llm.vhdl Comments URL: https://news.ycombinator.com/item?id=49957671 Points: 2 # Comments: 0
Halo is privacy focused personal assistant for iOS that uses a custom built agent harness with support to use you existing LLM subscriptions, Wiki based memory system backed by on-device RAG, a full-fledged local browser agent, chat with generative UI and more. TLDR: It's a better version of Hermes/OpenClaw for iOS, without needing a server! The better pa…
Article URL: https://www.harborframework.com/ Comments URL: https://news.ycombinator.com/item?id=49971276 Points: 1 # Comments: 0
Amazon SageMaker optimized generative AI inference introduces the aws-ai-ml skill through the Agent Toolkit for AWS, giving coding agents like Kiro, Claude Code, and Codex deep expertise in inference optimization and benchmarking. Describe what you want, and your agent generates executable SageMaker Python SDK v3 code to benchmark, recommend, and compare…
Article URL: https://brnch.io Comments URL: https://news.ycombinator.com/item?id=49966997 Points: 1 # Comments: 1
Article URL: https://github.com/Rikinshah787/dotpals Comments URL: https://news.ycombinator.com/item?id=49968767 Points: 1 # Comments: 1