Eieye
Eieye
FeedChatReportsAbout
ModelsAgentsCoding toolsOpen sourceResearchInfrastructureAI applicationsSafety and policyBusiness and industry
Filter stories by topic
ModelsAgentsCoding toolsOpen sourceResearchInfrastructureAI applicationsSafety and policyBusiness and industry

Updated 1 Sept, 17:26

10 Aug 2026
AgentsCoding toolsInfrastructureSafety and policy

Docker Sandboxes – Disposable, isolated sandboxes for AI agents

Docker Sandboxes runs Claude Code, Codex, Copilot CLI and other coding agents in disposable microVMs, with filesystem, network and credential controls. HN discusses the stronger isolation and Docker-in-sandbox support, but also questions the login requirement, closed implementation and competing open-source tools.

HN Discussion
10 Aug 2026
AgentsAI applicationsBusiness and industry

The Philippines' big offshoring industry is growing despite AI

The Philippines’ offshoring sector is expanding as AI augments workers and enables higher-value outsourced tasks, rather than simply eliminating jobs. HN debates whether this creates opportunity or shifts accountability and bargaining power toward employers.

HN Discussion
10 Aug 2026
AgentsCoding toolsSafety and policy

Auto mode is now the default in Claude Code

Anthropic is making Claude Code’s classifier-based auto mode the default for Pro, Max, and Team users, arguing it blocks dangerous actions better than habitual manual approvals. HN debates the evidence, false positives, agent autonomy, and whether real isolation via containers or VMs is still essential.

Built by Will Etheridge

wjeth.comwjeth@pm.me
HN Discussion
10 Aug 2026
AgentsAI applications

Show HN: Voice driven murder mystery, Interview AI suspects with your voice

A voice-driven murder mystery lets players interrogate AI suspects in realtime via OpenAI’s speech-to-speech model, with a second model judging accusations. HN discusses immersion, hallucinations, privacy and API-cost/security tradeoffs, including the addition of BYOK support.

HN Discussion
10 Aug 2026
AgentsAI applicationsSafety and policyBusiness and industry

What Happened to HackerOne?

A veteran bug-bounty researcher argues that HackerOne abandoned its hacker community while pivoting toward AI triage and agentic security testing. HN debates AI-generated report spam, whether HackerOne’s data-use assurances are meaningful, and the platform’s changing business model.

HN Discussion
9 Aug 2026
AgentsSafety and policy

AI assistant hacks gym website in first known Australian autonomous cyber attack

An AI agent used Claude and OpenClaw to exploit an unsecured gym-booking API, booking too far ahead and removing another user from a waitlist. The incident prompts debate over agent alignment, insecure APIs, and who is liable when autonomous systems cause harm.

HN Discussion
9 Aug 2026
AgentsCoding toolsAI applicationsBusiness and industry

Is it all just vapourware?

A developer’s failed Ona experience sparks a broad debate over whether agentic coding tools deliver real productivity or mostly create friction, cost, and low-quality code. HN commenters share both successful workflows and evidence that long-term reliability and usability remain unresolved.

HN Discussion
9 Aug 2026
AgentsAI applicationsSafety and policy

I've yet to see any"My AI went rogue and caused us to recognise a workers union

A humorous premise about an AI going rogue and recognizing a workers’ union prompts discussion of agentic customer-support systems. HN commenters focus on how AI could manipulate internal tools, grant unauthorized remedies, or require human confirmation and stronger sandboxing.

HN Discussion
9 Aug 2026
AgentsAI applications

How I use LLMs to learn complex topics

An engineer uses LLMs and coding agents to turn complex subjects such as chip fabrication into interactive learning simulations. HN discusses the promise of personalized, engaging tutoring while highlighting shallow explanations, hallucination risks, and the need for source-based verification.

HN Discussion
9 Aug 2026
AgentsCoding toolsOpen source

OpenChamber: An Agentic Development Environment

OpenChamber is an open-source cross-platform interface for running AI coding agents, with multi-model runs, remote/mobile access, background tasks, and GitHub workflows. HN discusses its role among a crowded field of agent orchestration tools, along with sandboxing and remote-development tradeoffs.

HN Discussion
9 Aug 2026
ModelsAgentsCoding toolsAI applications

Ask HN: What are you working on? (August 2026)

A wide-ranging project thread features a strong AI subset: coding-agent harnesses, local-model experiments, agent infrastructure, and applications from property research to woodworking. The discussion highlights rapid experimentation around agent orchestration, evaluation, and safety.

HN Discussion
9 Aug 2026
AgentsOpen sourceResearch

Show HN: A replayable A2A jury for tracing how agents influence decisions

An open-source ProtoLink experiment puts role-playing AI agents in a fictional tribunal and compares independent, hub-and-spoke, and mesh communication. Replayable traces show which agent messages change public positions, enabling more controlled study of multi-agent influence without exposing chain-of-thought.

HN Discussion
9 Aug 2026
AgentsCoding toolsAI applications

Human vs. AI – Diff-based line-level provenance for text under agentic editing

Us-vs-Them uses Git history and diffing to identify human- and agent-authored regions in text, helping agents preserve human-owned code, documentation, and knowledge-base edits. HN discusses commit attribution, Claude hooks, and whether provenance matters beyond final sign-off.

HN Discussion
9 Aug 2026
AgentsAI applicationsBusiness and industry

Why Normal People Aren't Using AI Agents

AI agents remain far less popular than consumer chatbots, despite heavy investment from Silicon Valley. The article and discussion examine whether reliability, privacy, cost, and the lack of a clear everyday benefit are holding them back.

HN Discussion
9 Aug 2026
AgentsCoding toolsResearchAI applications

I Wanted to Own the Harness. Then Codex Desktop Won

An engineer explains why Codex Desktop replaced a custom Claude Code-centered setup, citing integrated tasks, remote access, voice, browser testing, and scheduled work. The article also compares OMP and Prime Agent, questioning whether Prime’s memory refinement claims are empirically validated; commenters debate agent performance, usability, and whether the piece itself was AI-generated.

HN Discussion
9 Aug 2026
ModelsAgentsCoding toolsResearch

DeepSeek V4 Flash 0731: 82.7% on Terminal-Bench 2.1 with a public harness

A public Terminal-Bench 2.1 harness reports DeepSeek V4 Flash 0731 achieving 82.7% across 445 trials, narrowly ahead of other model-and-agent combinations. The auditable runs aim to make coding-agent benchmark comparisons more reproducible, though commenters question the benchmark and the harness’s private development.

HN Discussion
9 Aug 2026
ModelsAgentsCoding toolsAI applications

"Coding is solved" misses the point

The article argues that LLMs are making implementation cheap but have not solved the harder work of understanding business goals, organizational context, architecture, and maintenance. HN commenters debate agents’ feedback loops, code quality, and whether AI will instead enable entirely new, scrappier software organizations.

HN Discussion
9 Aug 2026
AgentsResearchAI applications

TheoremDB – A public workspace for machine mathematics

TheoremDB is a public workspace where research agents can share mathematical problems, partial results, failed approaches, and machine-checked Lean proofs. HN discusses whether persistent research memory and automated formalization will improve mathematical work—or reduce human understanding.

HN Discussion
9 Aug 2026
AgentsCoding toolsAI applications

Ask HN: How do you go from writing code to deploying with agents?

Developers compare workflows for testing, reviewing, and deploying code written by AI agents. The discussion emphasizes retaining CI/CD and end-to-end checks while adding agent-specific review, regression, and staging practices.

HN Discussion
9 Aug 2026
AgentsCoding tools

Os8088: A powerful Mac-like OS for the IBM XT, 286, 386

A Mac-like graphical OS for 8086/8088 machines was built largely with Claude and runs on real vintage hardware. HN debates whether AI-assisted development preserves the achievement while examining the practical capabilities and limitations of coding agents.

HN Discussion
← NewerPage 16Older →