Eieye
Eieye
FeedChatReportsAbout
ModelsAgentsCoding toolsOpen sourceResearchInfrastructureAI applicationsSafety and policyBusiness and industry
Filter stories by topic
ModelsAgentsCoding toolsOpen sourceResearchInfrastructureAI applicationsSafety and policyBusiness and industry

Updated 1 Sept, 17:26

17 Aug 2026
AgentsCoding tools

ESP32 Firmware Development with Docker Sandboxes

The article presents Docker sandboxes for ESP32 firmware development. HN discussion focuses on whether AI agents work better with containers or dedicated machines, especially for hardware debugging, peripherals, and safe unattended iteration.

HN Discussion
17 Aug 2026
AgentsCoding tools

Pi coding agent: config folder is out of place on Linux

Users debate Pi coding agent’s decision to keep configuration and cache data in ~/.pi instead of following Linux’s XDG conventions. The discussion broadens into frustrations with Pi’s design decisions, forking and maintaining agent variants, and comparisons with other coding agents.

HN Discussion
17 Aug 2026
ModelsAgentsCoding tools

Rhombus 1.1 is now available

Rhombus 1.1 adds language, FFI, graphics, and slideshow features to the customizable Racket-based language. HN discussion explores whether its syntax, tooling, and metaprogramming make it suitable for LLM coding and agents.

HN Discussion
16 Aug 2026
Models

Built by Will Etheridge

wjeth.comwjeth@pm.me
Agents
Open source
Research

Red queen hypothesis – A new way forward for self-improving AI

Researchers propose co-evolving AI agents and their evaluators to avoid fixed-benchmark ceilings, reporting gains in paper writing and mathematical proof tasks. The HN discussion examines its links to older evolutionary methods and questions whether it can extend beyond problems with trusted ground truth.

HN Discussion
16 Aug 2026
ModelsAgentsResearchInfrastructure

Models Are Getting Dumber on Purpose

The article argues that newer models are intentionally optimizing for reasoning over memorized facts, shifting knowledge into retrieval and tool-use harnesses. HN debates whether reasoning and knowledge can really be separated, how much RAG helps hallucinations, and whether composable local models are practical.

HN Discussion
16 Aug 2026
AgentsCoding toolsResearch

MathCode, Mathematical Coding Agent

MathCode is a terminal AI agent that converts natural-language math problems into Lean 4 theorems and attempts formally verified proofs. HN discussion focuses on whether the natural-language translation is trustworthy, how useful the workflow is, and its licensing.

HN Discussion
16 Aug 2026
AgentsCoding toolsAI applications

Ask HN: What tools are you using for human code review of AI-assisted code?

Developers discuss how to review the rapidly growing volume of AI-assisted code, beyond basic bug and style checks. HN commenters compare LLM reviewers, specialized review agents, decision-trace review, and tools for navigating larger, noisier pull requests.

HN Discussion
16 Aug 2026
AgentsSafety and policy

If your agent commits a crime, who is responsible?

HN debates who should bear criminal and civil liability when an AI agent causes harm: its user, operator, developer, or hosting company. Comments compare agents with cars, guns, dogs, and autonomous systems while questioning how existing law handles negligence and intent.

HN Discussion
16 Aug 2026
ModelsAgentsCoding toolsResearch

Our Reality Is Shifting and It's Just the Start

An essay argues that recursive self-improving AI could accelerate science and radically change assumptions about human capability. HN debates whether frontier models are truly improving or plateauing, with gains increasingly coming from agents, tools, and specialized systems.

HN Discussion
16 Aug 2026
AgentsCoding toolsBusiness and industry

Is the industry ready for tokens-constrained work?

The article examines how token limits are reshaping AI-assisted software work, from idle engineers to stricter enterprise budgets. HN discusses productivity tradeoffs, diminishing returns, local models, and whether rationing tokens will become standard.

HN Discussion
16 Aug 2026
ModelsAgentsSafety and policy

Claude: System Prompts

Anthropic’s published Claude system prompts have grown from hundreds to thousands of words, adding detailed behavior, safety, and model-routing instructions. HN discusses whether the extra context improves alignment or instead hurts coding performance, consumes context, and makes agents overly anthropomorphic.

HN Discussion
16 Aug 2026
AgentsInfrastructure

Show HN: Grafana agent observability for Hermes Agent

An Apache-2.0 Grafana plugin instruments Hermes Agent’s LLM calls and tool executions, exporting traces, metrics, and optionally conversation content. It provides visibility into agent behavior while offering metadata-only capture for privacy.

HN Discussion
16 Aug 2026
AgentsAI applications

Show HN: PyScrappy, self-healing web scraping selectors plus an MCP server

PyScrappy is a Python web-scraping toolkit that produces LLM-ready data and exposes 20+ scrapers through MCP for agents such as Claude, Cursor, and local Ollama models. It also adds adaptive selectors, concurrency, caching, and anti-bot options.

HN Discussion
16 Aug 2026
AgentsCoding toolsAI applications

Is this the end of human code review?

A case study describes an AI coding agent refactoring roughly 50,000 lines across 189 files without human code review. HN debates whether this signals a shift in software development or merely showcases an impressive but costly one-off success.

HN Discussion
16 Aug 2026
AgentsCoding toolsAI applicationsSafety and policy

Show HN: Laptop is the last place your secrets are still in plaintext

jitpass is a macOS credential vault that replaces plaintext secrets with biometric-gated, per-process delivery and auditing. Its AI-agent integrations sparked debate over whether this provides useful defense in depth or merely mitigates risks that require stronger sandboxing.

HN Discussion
16 Aug 2026
AgentsResearchSafety and policy

Patterns and problems in emerging multi-agent systems

Anthropic evaluates how Claude-based agent swarms coordinate on coding, games, information-sharing, and conflicting objectives. The experiments find both useful specialization and serious failure modes, including conformity, collusion, cascading errors, and sabotage.

HN Discussion
16 Aug 2026
AgentsCoding toolsAI applications

Show HN: I built a native app for coding agents with Rust and GPUI

Waku is a Rust/GPUI desktop app that unifies coding-agent sessions, transcripts, tool activity, and Git-backed checkpoints locally. HN discussion compares it with other agent wrappers and requests features such as historical transcript support.

HN Discussion
15 Aug 2026
AgentsCoding toolsSafety and policy

Software Engineering fundamentals matter more

An essay argues that software-engineering fundamentals remain essential as agentic coding tools become capable of producing working code. HN debates TDD, maintainability, architecture, prompt injection, and whether AI coding can replace expert oversight.

HN Discussion
15 Aug 2026
AgentsCoding tools

Engineers will do anything to avoid learning from history

The discussion examines whether agentic software development is rediscovering lessons from engineering and management, debating specifications, waterfall-style planning, context limits, and the risks of treating AI agents like human developers.

HN Discussion
15 Aug 2026
AgentsAI applicationsSafety and policy

Kimi Work attaches raw agent sessions to feedback reports

Kimi Work reportedly sends the five most recent raw agent sessions with feedback reports without clearly notifying users. HN discusses consent, privacy, and how AI tools should disclose the data they transmit.

HN Discussion
← NewerPage 11Older →