Eieye
Eieye
FeedChatReportsAbout
ModelsAgentsCoding toolsOpen sourceResearchInfrastructureAI applicationsSafety and policyBusiness and industry
Filter stories by topic
ModelsAgentsCoding toolsOpen sourceResearchInfrastructureAI applicationsSafety and policyBusiness and industry

Updated 1 Sept, 17:26

29 Aug 2026
AgentsResearchAI applications

I accidentally turned LLM memory into program analysis

Lemmalog gives LLM agents a maintained Datalog state instead of relying on transcript retrieval, enabling provenance, retractions, and temporal updates. Its benchmarks show substantially smaller query context and stronger handling of knowledge updates, though extraction and inference remain weak.

HN Discussion
28 Aug 2026
ResearchAI applicationsSafety and policy

Identifying fake cosmetics using AI

A lab tests Gemini’s image analysis on authentic and counterfeit Rhode lip tints. It catches subtle packaging errors but also mistakes glare for defects and falsely labels a genuine product counterfeit, prompting debate about AI’s reliability and overconfidence.

HN Discussion
28 Aug 2026
ModelsResearch

Racter (1984)

Racter was a 1984 BASIC program that generated seemingly coherent English prose from rules, variables, and randomized selections. HN discusses it as an early ancestor of modern language models and conversational AI.

HN Discussion
28 Aug 2026

Built by Will Etheridge

wjeth.comwjeth@pm.me
Agents
Open source
Infrastructure
Safety and policy

Show HN: Conduct, open-source guardrails for LLM and MCP tool calls

Conduct is an open-source control plane for governing AI agents across LLM, shell, and MCP tool calls. It combines policy enforcement, signed configurations, and hash-chained audits, while HN commenters point to lighter or alternative sandboxing approaches.

HN Discussion
28 Aug 2026
AgentsCoding toolsAI applications

Aspirational Clownmaxxing and Joey's cadillac todo list

An experiment pushes Claude coding agents with surreal prompts to build elaborate PySide6 todo apps, finding that the outputs mostly collapse into familiar AI design patterns. HN discusses model choice, prompting, creative limits, and the persistence of AI-generated slop.

HN Discussion
28 Aug 2026
AI applicationsSafety and policy

Show HN: Sesame - a local-first, open-source password manager

HN commenters focus heavily on claims that Sesame was vibe-coded with AI, questioning whether LLM-generated code and UI are appropriate for security-critical password management. The broader discussion also examines password-manager threat models, malware resistance, and usability trade-offs.

HN Discussion
28 Aug 2026
ResearchAI applications

The Analytical AI Handbook

The Analytical AI Handbook covers using foundation models to transform unstructured data into reliable decisions through classification, extraction, judging, and evaluation. It emphasizes measurable tasks, smaller models, batch processing, and production architectures.

HN Discussion
28 Aug 2026
Coding toolsBusiness and industry

Ask HN: AI writes better code than me. How to keep my identity?

A freelancer describes losing motivation and professional identity as Claude and GPT increasingly outperform him on routine coding. HN discusses shifting value toward architecture, domain expertise, reviewing AI output, and broader career changes.

HN Discussion
28 Aug 2026
ModelsResearch

Separating logic and language

An MIT-led study finds that severe language impairment does not prevent logical reasoning, suggesting distinct brain systems for language and logic. HN discussion explores whether this supports separating language interfaces from reasoning and control in LLM-based systems.

HN Discussion
28 Aug 2026
AI applicationsSafety and policy

LibreOffice 26.8 is out – local first, and with no AI

LibreOffice 26.8 emphasizes local-first use and explicitly rejects built-in AI, while adding language, accessibility, and formatting improvements. The article frames its AI-free stance as a deliberate alternative to cloud office suites integrating AI.

HN Discussion
28 Aug 2026
InfrastructureBusiness and industry

Data centers' 'oh s–t' moment

HN discusses whether mounting opposition to AI data-center construction could constrain GPU supply and expose a broader AI investment bubble. Comments debate local environmental concerns, financing risks, and the economic impact of slower infrastructure growth.

HN Discussion
28 Aug 2026
AgentsCoding toolsAI applications

How I Design with AI

An exploration of using AI agents and Claude for web and product design, with commenters discussing constraint-based workflows, design systems, and the persistent problem of generic or inaccurate AI-generated interfaces.

HN Discussion
28 Aug 2026
ModelsAgentsOpen sourceResearch

Autonomous Mathematical Discovery in an Open-World Multi-Agent Environment

An open-world multi-agent system reportedly discovered novel mathematical constructions and theorems across several research problems without a central coordinator. The paper releases agent dialogues, proofs, verification artifacts, and code, prompting debate about AI creativity and the future of mathematical research.

HN Discussion
28 Aug 2026
AgentsOpen sourceResearchSafety and policy

Just the rumour of a bug is enough to find an exploit these days

An OCaml maintainer argues that AI agents can turn vague vulnerability reports into working exploits within minutes, undermining traditional security embargoes. HN discusses automated patch analysis, the race between attackers and maintainers, and possible defenses such as faster releases and virtual patching.

HN Discussion
28 Aug 2026
InfrastructureBusiness and industry

Nvidia Insists It Can Keep Printing Money to Fund the AI Boom

Nvidia argues it can keep funding massive AI-related investments, including infrastructure projects, from its exceptional cash flow. HN debates whether AI demand is genuine or subsidized, and whether custom chips could eventually challenge Nvidia’s dominance.

HN Discussion
28 Aug 2026
AgentsCoding toolsAI applications

How Dactyl Works

Dactyl uses a WebAssembly SwiftUI renderer to let AI generate and preview native-style apps across iOS, Android, and the web. Its development loop uses visual LLMs and coding agents to compare output with Apple’s simulator and improve compatibility.

HN Discussion
28 Aug 2026
ModelsInfrastructureAI applications

Run Qwen3.8 27B locally: real numbers from my Mac Studio

A hands-on benchmark of Qwen3.8 27B on Apple silicon compares quantizations, runtimes, speed, memory needs, and practical background-assistant workloads. HN commenters question the unusually slow results and discuss optimizations, GPUs, and the privacy benefits of local inference.

HN Discussion
28 Aug 2026
ModelsCoding toolsOpen sourceResearch

GLM-5.3 is now open-weight

Z.ai has released GLM-5.3 as an open-weight model, claiming major post-training gains in coding, long-horizon agents, and cybersecurity. HN users discuss real-world performance, hosting costs, quantization, local deployment, and the risks of its cyber capabilities.

HN Discussion
28 Aug 2026
AgentsCoding toolsAI applications

Superhuman Attention

Perfloop uses AI agents to discover, implement, verify, and repeatedly measure performance improvements, aiming to keep machine-generated code from overwhelming human reviewers. The article argues that automated proof should preserve scarce engineering attention.

HN Discussion
28 Aug 2026
AgentsCoding toolsInfrastructureAI applications

The Finn – an agent that lives in my router and complains about it

The Finn is a small LLM-powered network-monitoring agent that runs locally on an OpenWrt router, using model calls only when it detects unusual activity. The project explores autonomous, constrained agents with a physical vantage point, while the discussion questions its cloud-model dependency and usefulness.

HN Discussion
← NewerPage 6Older →