Eieye
Eieye
FeedChatReportsAbout
ModelsAgentsCoding toolsOpen sourceResearchInfrastructureAI applicationsSafety and policyBusiness and industry
Filter stories by topic
ModelsAgentsCoding toolsOpen sourceResearchInfrastructureAI applicationsSafety and policyBusiness and industry

Updated 3 Sept, 07:44

14 Aug 2026
ModelsAgentsOpen source

Qwen3.8 27B

Qwen releases Qwen3.8-27B, a 27B open vision-language model supporting configurable reasoning, 262K-native context, video understanding, and agentic coding. HN users discuss its impressive benchmark gains and the challenge of running it on consumer hardware.

HN Discussion
14 Aug 2026
AI applicationsSafety and policy

Error by AI scribe during medical appointment leaves patient devastated

An AI medical scribe added a false claim of psychedelic drug use to a patient’s record, highlighting risks from unchecked hallucinations. HN discusses whether these systems reduce errors, how speech-recognition failures arise, and who is accountable when clinicians skip review.

HN Discussion
14 Aug 2026
Open sourceInfrastructureAI applications

Show HN: Lumabri – Run Moe Models on a P2P Swarm with Colibri

Lumabri is an Apache-licensed P2P inference system that spreads large mixture-of-experts models across CPUs and GPUs, transferring only needed weights or activations. HN discussion highlights its LAN use cases, latency and throughput tradeoffs, and the difficulty of guaranteeing deterministic results across hardware.

HN Discussion

Built by Will Etheridge

wjeth.comwjeth@pm.me
14 Aug 2026
AI applications

SparrowMap – Cameras that watch government vehicles

SparrowMap uses on-device computer vision to identify government vehicles and license plates while discarding other plates locally. HN discussion debates privacy, safety, and whether crowdsourced sousveillance can counter commercial camera networks.

HN Discussion
14 Aug 2026
ResearchAI applicationsSafety and policy

How AI text watermarking works

An interactive explainer shows how providers watermark AI text by subtly biasing token choices, allowing secret-key statistical detection while surviving copying. HN debates whether the marks can identify users, withstand rewriting, work on code, prevent model collapse, or become a regulatory and commercial tool.

HN Discussion
13 Aug 2026
AgentsCoding toolsInfrastructure

Show HN: MCP-stama – An ultra-fast Rust MCP server with no dependencies

mcp-stama is a dependency-free Rust MCP server providing fast file search, Git, and Docker tools to AI coding agents. It emphasizes sub-millisecond responses, low memory use, and simple editor configuration.

HN Discussion
13 Aug 2026
ModelsOpen sourceInfrastructureAI applications

Chestnut – eGPU dock with open-source firmware

Comma.ai’s Chestnut is an open-source-firmware USB4 eGPU dock designed to run substantially larger openpilot driving models in cars. HN discusses its tinygrad integration, hardware limitations, cost, and safety fallback behavior.

HN Discussion
13 Aug 2026
Infrastructure

Terabytes of credentials leaked in supply-chain attack

A supply-chain compromise of LiteLLM exposed credentials from roughly 2,500 organizations and 434,000 CI/CD pipelines. The breach highlights the systemic risk of rushing AI tooling into software supply chains and failing to rotate secrets.

HN Discussion
13 Aug 2026
ResearchAI applicationsBusiness and industry

How Organizations Use AI: Evidence from ChatGPT [pdf]

An OpenAI-affiliated economics paper examines how organizations use ChatGPT, including adoption patterns and differences across workers and industries. HN debates its methodology, weak evidence for ROI, and whether the paper reads more like marketing than rigorous analysis.

HN Discussion
13 Aug 2026
ModelsCoding toolsBusiness and industry

Ask HN: How much money do you spend monthly on subscriptions for AI models?

HN users compare monthly spending on Claude, ChatGPT, Gemini, coding assistants, APIs, and local models. The discussion highlights wide cost variation, from free and local setups to hundreds or thousands of dollars, along with debates over subscription value and usage limits.

HN Discussion
13 Aug 2026
ResearchAI applicationsSafety and policy

Person Hides Prompt Injection in Legal Filing Telling AI to Side with Them

A self-represented litigant hid prompt injections in court filings to make any reviewing AI favor his case, prompting sanctions and warnings from the judge. The incident highlights risks for AI-assisted legal workflows and the difficulty of preserving document integrity.

HN Discussion
13 Aug 2026
AgentsCoding toolsAI applications

Understanding is the new bottleneck

A talk argues that coding agents shift the bottleneck from writing code to understanding it, proposing explainers, quizzes, interactive “micro-worlds,” and shared workspaces to keep humans engaged. HN debates whether these techniques can offset the complexity, technical debt, and review burden created by agent-generated code.

HN Discussion
13 Aug 2026
InfrastructureSafety and policy

AI Is Threatening Natural Resources for Billions

A UN-linked report warns that AI’s rapidly expanding data-center footprint is putting pressure on water, energy, land, and minerals. HN debates the scale of those costs, comparisons with other industries, and whether stronger measurement and regulation are needed.

HN Discussion
13 Aug 2026
ModelsAgentsInfrastructureBusiness and industry

Accelerating GPT-5.6 Sol Ultrafast

OpenAI and Cerebras are previewing Ultrafast, a GPT-5.6 Sol API tier delivering up to 750 output tokens per second. HN debates its benchmark claims, likely premium pricing, Cerebras’s wafer-scale architecture, and how low-latency inference could change agentic coding and real-time applications.

HN Discussion
13 Aug 2026
AgentsCoding toolsResearch

How Compaction Works in Pi

Pi explains how its coding agent compacts long LLM conversations using a separate summarization request while retaining recent turns. HN discusses the trade-offs with pruning, prompt caching, latency, context loss, and alternative approaches such as image-based or server-side compaction.

HN Discussion
13 Aug 2026
AI applications

Solid 2.0 RC: The Big <Reveal>

Solid 2.0 introduces first-class async primitives and consolidates framework capabilities, with commenters praising its developer experience and questioning bundle-size tradeoffs. Much of the discussion focuses on the announcement’s heavily AI-assisted prose and whether it obscures the technical advances.

HN Discussion
13 Aug 2026
ModelsSafety and policyBusiness and industry

Amazon will train on Twitch streamers' content by default, unless they opt out

Twitch will let Amazon use creators’ livestreams to train generative AI models unless they opt out, prompting backlash over consent and transparency. HN commenters focus on the deliberately non-consensual default and Twitch’s uncertainty about whether training has already occurred.

HN Discussion
13 Aug 2026
ModelsAgentsCoding toolsBusiness and industry

Gemini 3.7 Flash

Google’s Gemini 3.7 Flash targets coding and agent workflows with reported benchmark gains, faster responses, and introductory API pricing of $0.75 per million input tokens. HN users debate its real-world quality, strong multimodal performance, competitiveness against cheaper models, and Google’s confusing API and billing experience.

HN Discussion
13 Aug 2026
ModelsAI applicationsBusiness and industry

Mistral OCR 4.1

Mistral OCR 4.1 is discussed as a faster, specialized document-understanding model with layout and bounding-box extraction. HN users debate its accuracy, hallucinations, privacy, and steep pricing versus local and competing OCR tools.

HN Discussion
13 Aug 2026
ModelsAgentsCoding toolsSafety and policy

Gemini 3.7 Flash

Google introduces Gemini 3.7 Flash, a cheaper model aimed at coding, complex workflows, and autonomous agents. It claims major benchmark gains over 3.6 Flash and adds updated cyber and CBRN safeguards.

HN Discussion
← NewerPage 37Older →