Eieye
Eieye
FeedChatReportsAbout
ModelsAgentsCoding toolsOpen sourceResearchInfrastructureAI applicationsSafety and policyBusiness and industry
Filter stories by topic
ModelsAgentsCoding toolsOpen sourceResearchInfrastructureAI applicationsSafety and policyBusiness and industry

Updated 1 Sept, 17:26

21 Aug 2026
ModelsInfrastructureSafety and policy

Ox Alpha

OpenRouter’s anonymous Ox Alpha model became a large-scale community investigation, with testing suggesting it is ZAI’s GLM-5.3 Flash. Discussion covers its coding and reasoning quality, censorship fingerprints, uncertain provider identity, and privacy trade-offs.

HN Discussion
20 Aug 2026
ModelsSafety and policyBusiness and industry

Aaron Swartz was prosecuted for scraping, while Meta does it without consequence

The post contrasts Aaron Swartz’s aggressive prosecution for downloading academic papers with Meta’s relatively limited consequences for allegedly torrenting books to train AI models. HN debates the legal differences, disproportionate enforcement, copyright, and whether powerful AI companies effectively receive preferential treatment.

HN Discussion
20 Aug 2026
ModelsAI applicationsBusiness and industry

Our Servants Will Do That for Us

An essay argues that AGI could automate not only drudgery but also meaningful professions, from software engineering to art, because consumers often prefer cheap, convenient machine-generated outcomes. HN debates whether people truly value human interaction and how society would distribute resources in a post-work economy.

HN Discussion

Built by Will Etheridge

wjeth.comwjeth@pm.me
20 Aug 2026
ModelsResearch

Why aren't smart people happier? (2022)

An essay argues that AI’s rapid progress is largely confined to well-defined tasks, while happiness and wisdom involve poorly defined problems. HN mostly debates the relationship between IQ, emotional maturity, effort, and life satisfaction.

HN Discussion
20 Aug 2026
ModelsResearchSafety and policy

What Is Reasoning

The article explains reasoning traces as model-generated text routed through special channels, and examines how prompts, prefilling, and hidden scratchpads control reasoning effort and leakage. HN commenters note safety-filter concerns and compare the behavior to rubber-duck debugging.

HN Discussion
20 Aug 2026
ModelsCoding toolsResearchInfrastructure

AI at Home Part 2: Multi-GPU Drifting

A detailed experiment optimizes Gemma, DeepSeek, and Qwen inference across four inexpensive AMD GPUs. Fixing PCIe peer-to-peer transfers and using tensor or layer parallelism substantially improves local LLM performance, making agentic coding workloads practical.

HN Discussion
20 Aug 2026
ModelsCoding toolsOpen sourceAI applications

Vomit: Clean up Claude 5's token output with a separate LLM

Vomit is a local LLM wrapper that rewrites Claude Code’s verbose, jargon-heavy output into clearer English through hooks. The discussion explores whether this extra model layer is worthwhile, alternative prompts and deterministic filters, and broader theories about Claude 5’s changed writing style.

HN Discussion
20 Aug 2026
ModelsResearchSafety and policy

Could AIs Become Conscious?

The story asks whether AI systems could develop genuine consciousness rather than merely simulate it. HN debates substrate, embodiment, qualia, agency, and whether uncertainty should shape how advanced models are trained and treated.

HN Discussion
20 Aug 2026
ModelsResearch

An elliptic curve of rank ≥ 30

A newly submitted elliptic curve has at least 30 independent rational points, breaking the previous rank record of 29. HN discussion focuses on the revelation that Claude, working with mathematicians, helped find the result and what this suggests about AI-assisted mathematical research.

HN Discussion
20 Aug 2026
ModelsResearchInfrastructureAI applications

A look under our trunk: what's in our compute

Waymo details the custom 5nm ASIC and heterogeneous, redundant onboard compute powering its autonomous Driver. The HN discussion examines the system’s hardware claims, safety and remote assistance, and the broader practicality of robotaxis.

HN Discussion
20 Aug 2026
ModelsBusiness and industry

Stwipe Acquires OpenWouter

A satire of Stripe’s OpenRouter acquisition imagines buying “OpenWouter,” a one-person AI model with a two-endpoint API. Its fake model card and benchmark claims lampoon AI product and M&A marketing.

HN Discussion
20 Aug 2026
ModelsResearchSafety and policy

Guess which of these LLM outputs is watermarked

An experiment suggests people cannot reliably distinguish watermarked LLM prose from unwatermarked text. The discussion explores SynthID’s sampling-based design, its limits on low-entropy text, and concerns about detection, paraphrasing, and intellectual-property implications.

HN Discussion
20 Aug 2026
ModelsResearchInfrastructure

DiffusionGemma Technical Report

DiffusionGemma fine-tunes Gemma 4 into a discrete-diffusion language model that generates roughly 1,500 tokens per second on an H100 by refining 256-token blocks in parallel. HN discusses its promising local-inference speed, implementation challenges, hardware tradeoffs, and quality gaps versus autoregressive models.

HN Discussion
20 Aug 2026
ModelsResearchInfrastructureAI applications

Show HN: I trained a 125M model to autocomplete piano on-device

An author trained a 125M-parameter transformer to continue live MIDI piano performances at about 108 notes per second on an iPhone. The write-up details compact note representations, data cleaning, DPO preference training, and Core ML deployment; HN discusses musical quality and extensions such as accompaniment and other instruments.

HN Discussion
20 Aug 2026
ModelsResearchAI applications

Seeing beyond BMI: Estimating cardiometabolic risk with smartphone imagery

PhotoScan uses deep learning on smartphone photos to estimate body-fat distribution and cardiometabolic risk, reporting near-DXA performance in small cohorts. HN debates validation limits, comparisons with prior systems, privacy, and possible insurance misuse.

HN Discussion
20 Aug 2026
ModelsOpen sourceInfrastructureBusiness and industry

If this is true, the hyperscalers are toast

An article argues that increasingly capable, energy-efficient small language models could move much AI inference from hyperscaler data centers to desktops and phones. HN debates benchmark limitations, local-model quality, privacy, economies of scale, and whether this really threatens frontier labs or cloud infrastructure.

HN Discussion
20 Aug 2026
ModelsResearch

Universality of Gradient Descent Neural Network Training

A theoretical result argues that any neural network with an algorithm capable of finding good weights can be extended so gradient descent reproduces them. The construction is impractical but informs the limits of meta-learning and neural-network optimization.

HN Discussion
19 Aug 2026
ModelsSafety and policyBusiness and industry

OpenAI's Unraveling Has Begun

The article argues that OpenAI’s frontier-training pause, rapidly growing losses, and opaque disclosures signal serious business and credibility problems ahead of a possible IPO. HN commenters dispute the author’s bearish interpretation while discussing OpenAI’s safety claims and the financial risks spreading through the AI infrastructure ecosystem.

HN Discussion
19 Aug 2026
ModelsResearchInfrastructure

DFlash 2: Keep Drafting Parallel

DFlash 2 improves parallel speculative decoding with a lightweight path selector and local convolution, reporting roughly 16–25% longer acceptance and up to 3× throughput. HN commenters test integrations, troubleshoot quantization, and discuss lossless sampling and stochastic outputs.

HN Discussion
19 Aug 2026
ModelsSafety and policy

Digital Immortality

An essay reflects on how human writing is absorbed into LLM training and what that means for identity, consent, and privacy. HN debates data poisoning, surveillance risks, and the unsettling idea of AI as a form of digital immortality.

HN Discussion
← NewerPage 7Older →