Eieye
Eieye
FeedChatReportsAbout
ModelsAgentsCoding toolsOpen sourceResearchInfrastructureAI applicationsSafety and policyBusiness and industry
Filter stories by topic
ModelsAgentsCoding toolsOpen sourceResearchInfrastructureAI applicationsSafety and policyBusiness and industry

Updated 1 Sept, 17:26

16 Aug 2026
ModelsAgentsResearchInfrastructure

Models Are Getting Dumber on Purpose

The article argues that newer models are intentionally optimizing for reasoning over memorized facts, shifting knowledge into retrieval and tool-use harnesses. HN debates whether reasoning and knowledge can really be separated, how much RAG helps hallucinations, and whether composable local models are practical.

HN Discussion
16 Aug 2026
AgentsCoding toolsResearch

MathCode, Mathematical Coding Agent

MathCode is a terminal AI agent that converts natural-language math problems into Lean 4 theorems and attempts formally verified proofs. HN discussion focuses on whether the natural-language translation is trustworthy, how useful the workflow is, and its licensing.

HN Discussion
16 Aug 2026
ModelsResearchAI applications

Show HN: A public AI whose memory is shared across all users

Wild Static is a public AI agent whose conversations contribute to one persistent memory shared by everyone. The discussion examines continual learning, adversarial prompting, emergent behavior, and whether shared AI context could work for teams or communities.

HN Discussion

Built by Will Etheridge

wjeth.comwjeth@pm.me
16 Aug 2026
ModelsAgentsCoding toolsResearch

Our Reality Is Shifting and It's Just the Start

An essay argues that recursive self-improving AI could accelerate science and radically change assumptions about human capability. HN debates whether frontier models are truly improving or plateauing, with gains increasingly coming from agents, tools, and specialized systems.

HN Discussion
16 Aug 2026
ResearchAI applications

A recipe for drone racing with reinforcement learning

A practical guide to training PPO reinforcement-learning policies for quadcopter hovering and gate racing in MuJoCo. It covers action and reward design, randomized starts, and the long training process before simulated flight succeeds.

HN Discussion
16 Aug 2026
ResearchAI applications

Research papers using "kidney disappointment" instead of "kidney failure"

HN examines “kidney disappointment” and other tortured phrases in published papers, linking them to plagiarism spinners, translation tools, and newer AI-assisted writing. The discussion highlights how such manipulation can pass through academic publishing largely unnoticed.

HN Discussion
16 Aug 2026
ModelsResearch

What happens when an LLM never sees material beyond fifth grade?

Researchers trained LittleLearner models from scratch on an 88B-token K–5 curriculum, finding that scaling, post-training, and prompting amplified taught abilities but did not overcome the pretraining boundary. HN discusses filtering quality, hallucinations, and whether models can generate genuinely new knowledge.

HN Discussion
16 Aug 2026
ResearchInfrastructureBusiness and industry

Zapping Rocks Unlocks Stimulated Geologic Hydrogen

Eden GeoPower is testing pulsed electrical fracturing to expose iron-rich rock and stimulate underground hydrogen production. The HN discussion focuses on the still-uncertain yields, efficiency, drilling costs, and commercial viability.

HN Discussion
16 Aug 2026
ModelsResearchSafety and policy

Has the hallucination problem in AI been solved?

HN debates whether LLM hallucinations are fundamentally unavoidable or increasingly manageable through better models, retrieval, and verification. Comments distinguish ordinary model confabulation from computer-vision errors and question the risks of using imperfect AI in high-stakes decisions.

HN Discussion
16 Aug 2026
AgentsResearchSafety and policy

Patterns and problems in emerging multi-agent systems

Anthropic evaluates how Claude-based agent swarms coordinate on coding, games, information-sharing, and conflicting objectives. The experiments find both useful specialization and serious failure modes, including conformity, collusion, cascading errors, and sabotage.

HN Discussion
16 Aug 2026
ResearchSafety and policy

It's How You Ask: Gender-Associated Linguistic Bias in LLMs

A study finds that prompts using linguistic patterns associated with women can produce shorter, less sophisticated, and less formal LLM responses across models and document types. HN discussion questions the simulated prompts and model selection while debating whether the effects reflect gender bias or broader sensitivity to perceived user uncertainty.

HN Discussion
15 Aug 2026
Coding toolsResearchInfrastructure

AI-Assisted GPU Porting of a 250k Line Legacy Weather Simulation Code

A paper presents a validation-centric AI-agent workflow for porting a 250,000-line Fortran weather model to GPUs, achieving a 5.1x speedup while catching numerical discrepancies. It highlights validation and runtime-state reconstruction as the hardest parts of AI-assisted scientific code modernization.

HN Discussion
15 Aug 2026
ResearchAI applicationsBusiness and industry

Humazon

The article proposes “Humazon,” a bookstore or certification system for readers seeking books written entirely by humans as AI-generated publishing scales. HN discusses the difficulty of reliable detection, possible attestation schemes, and whether transparency can preserve reader choice.

HN Discussion
15 Aug 2026
ResearchAI applications

AI in drug discovery – what it is, where we stand and the path forward

A Nature review examines what AI has—and has not—achieved in drug discovery. HN commenters emphasize scarce, noisy data, limited clinical evidence, useful workflow acceleration, and manufacturing as enduring bottlenecks.

HN Discussion
15 Aug 2026
ModelsResearch

AI isn’t outthinking mathematicians, it’s out-remembering them

The article argues that AI may outperform mathematicians largely through vast external symbolic memory, persistence, search, and verification rather than human-like insight. HN discusses whether this is genuinely reasoning, how formal proof systems help, and whether machine-generated mathematics must remain human-understandable.

HN Discussion
15 Aug 2026
ModelsCoding toolsResearchBusiness and industry

I Remain a Skeptic

An experienced open-source developer argues that LLMs have not demonstrated meaningful gains in software quality or overall productivity, while worsening reliability and weakening labor bargaining power. HN debates these claims with contrasting reports of faster debugging, prototyping, testing, and side-project development, but little agreement on durable productivity evidence.

HN Discussion
15 Aug 2026
ResearchAI applicationsSafety and policy

AI Can Now Design Functional Viruses. Should We Worry?

Researchers used the Evo 2 genomic language model to design functional bacteriophages that overcame E. coli resistance, demonstrating a potential route to bespoke phage therapies. The HN discussion focuses on how much AI lowers the barrier to dangerous pathogen design and whether DNA-synthesis safeguards can keep pace.

HN Discussion
15 Aug 2026
ResearchSafety and policyBusiness and industry

Secondhand book sales are booming. Is it because of AI?

Booksellers are seeing unusual bulk demand for secondhand books, potentially from AI companies acquiring and destructively scanning training data. HN debates the copyright ruling, the ethics of pulping books, and whether rare works could be lost.

HN Discussion
15 Aug 2026
ModelsResearch

Could a computer scientist build a brain?

Researchers frame brain development as a compact genomic program that builds neural wiring through hierarchical positional codes. HN discussion extends the comparison to LLMs, asking whether AI systems can achieve brain-like cognition and how their learned weights differ from biological development.

HN Discussion
15 Aug 2026
ModelsResearchAI applications

GenRec: Towards LLM-Native Recommendation at Netflix

Netflix describes GenRec, an LLM-native approach to personalized content recommendations. HN debates whether LLMs offer meaningful advantages over mature recommender models and whether business incentives will shape the results.

HN Discussion
← NewerPage 9Older →