Eieye
Eieye
FeedChatReportsAbout
ModelsAgentsCoding toolsOpen sourceResearchInfrastructureAI applicationsSafety and policyBusiness and industry
Filter stories by topic
ModelsAgentsCoding toolsOpen sourceResearchInfrastructureAI applicationsSafety and policyBusiness and industry

Updated 1 Sept, 17:26

10 Aug 2026
ModelsResearchSafety and policy

GPT 5.6 Cyber

GPT 5.6 Cyber is an access-restricted OpenAI model aimed at cybersecurity work. HN discusses its offensive-security capabilities, inconsistent guardrails, identity verification, and whether restricting access meaningfully improves safety.

HN Discussion
10 Aug 2026
ModelsAI applications

Show HN: Vocal Slice – Cut audio by selecting text, fully on-device

Vocal Slice is a desktop audio editor that uses local Whisper transcription to cut recordings by highlighting words, with precise waveform adjustment and offline, privacy-preserving processing. HN discusses its workflow advantages, accuracy, and how readily similar tools can now be built with open models and coding agents.

HN Discussion
10 Aug 2026
ModelsAI applications

Itadakimasu: A word you say to the food, not the cook

The article examines who or what “itadakimasu” addresses before a meal, alongside the related post-meal phrase “gochisousama.” HN’s discussion focuses heavily on whether its polished, over-intellectualized prose was AI-generated and what that means for trust in personal writing.

HN Discussion
10 Aug 2026

Built by Will Etheridge

wjeth.comwjeth@pm.me
ModelsResearch

Exploring Claude/GPT Knowledge Cutoffs and Pre-Training Timelines

A probing methodology estimates frontier models’ knowledge cutoffs, training timelines, data mixtures, and exposure to other models’ outputs. HN discusses how fixed API weights, post-training, distillation, and product-layer updates complicate those inferences.

HN Discussion
10 Aug 2026
ModelsOpen sourceSafety and policyBusiness and industry

Mark Zuckerberg attacks 'closed' AI rivals as Meta returns to open models

Mark Zuckerberg argues that AI should remain broadly accessible as Meta resumes releasing open-weight models and attacks closed rivals. HN debates whether this expands competition and user control or is mainly a strategic effort to commoditize competitors, alongside concerns about safety and what “open” really means.

HN Discussion
10 Aug 2026
ModelsInfrastructureBusiness and industry

AI's profits are 'being funded by investors rather than earned from customers

An analysis argues that AI’s upstream profits are being funded by investor capital while model and application companies remain deeply unprofitable. HN commenters debate whether acquisitions, debt, and continued infrastructure spending can sustain the boom.

HN Discussion
10 Aug 2026
ModelsAgentsResearch

Humanising LLM Outputs Is Dumb

The article argues that forcing LLMs and coding agents into concise, human-friendly styles during task execution can lose useful detail and hide uncertainty. HN debates whether style instructions actually harm reasoning, with many favoring structured machine-facing state followed by a final human-oriented rendering step.

HN Discussion
10 Aug 2026
ModelsCoding toolsInfrastructureAI applications

DeepSeek costs OpenCode Go user $1.14/day; dual DGX breaks even in 24 years

A comparison finds that an average OpenCode Go user’s DeepSeek usage costs about $1.14 per day, making a $10,000 dual-DGX setup take decades to break even. HN discusses workload assumptions, local-model privacy and performance, cloud economics, and OpenCode’s data-handling concerns.

HN Discussion
10 Aug 2026
ModelsOpen sourceResearchInfrastructure

Show HN: A tiny LLM running at 21,000 tok/s on a $250 FPGA (Live Demo)

An open-source project runs a tiny 3.16M-parameter INT4 language model entirely in a $250 FPGA’s on-chip SRAM, reaching about 21,000 tok/s in a usable single-stream build. The demo is intentionally impractical as a chatbot, but illustrates how eliminating DRAM traffic can transform inference speed and why larger models remain difficult.

HN Discussion
10 Aug 2026
ModelsSafety and policyBusiness and industry

A 'bananas' order for 5000 obscure book titles fuels suspicion

European booksellers suspect automated bulk orders are sourcing books for AI training, potentially through destructive scanning and cross-border workarounds. HN debates copyright, preservation, fair use, and whether the practice harms or supports the used-book industry.

HN Discussion
10 Aug 2026
ModelsResearch

I Benchmarked Local LLMs on the Laptop I Have

A benchmark examines how local LLMs perform on an ordinary laptop. HN discussion weighs very low local throughput and hardware limits against cheap hosted models, while sharing newer model variants to test.

HN Discussion
10 Aug 2026
ModelsAI applicationsBusiness and industry

ChatGPT Knows Who It'll Recommend Before It Searches

An experiment argues that ChatGPT often inserts brands into its initial search query before retrieving results, strongly favoring products it already associates with a category. The small, single-account study suggests AI visibility depends both on being in that shortlist and on having pages likely to be cited.

HN Discussion
10 Aug 2026
ModelsAgentsOpen sourceSafety and policy

Meta's new open-weight model targets local agentic AI

Meta is releasing an open-weight model aimed at running agentic AI locally. HN discussion focuses on the trade-offs of open access, containment risks, and skepticism about Meta’s motivations.

HN Discussion
10 Aug 2026
ModelsAgentsOpen sourceInfrastructure

Muse Glimmer: 30B-parameter model optimized for always-on local agent workflows

Meta released Muse Glimmer, a permissively licensed 30B multimodal model designed for local agent workflows on consumer hardware. HN users test its coding and tool-use quality, debate benchmark claims, and examine its quantization, speed, memory needs, and advantages over hosted APIs.

HN Discussion
9 Aug 2026
ModelsOpen sourceInfrastructureSafety and policy

Show HN: Lumabri – What if LLMs worked like Napster?

Lumabri is an open-source peer-to-peer system for running huge mixture-of-experts LLMs across ordinary CPUs, disks, and GPUs. HN discussion focuses on whether network latency, untrusted computation, privacy, and abuse safeguards make the approach practical.

HN Discussion
9 Aug 2026
ModelsAgentsCoding toolsAI applications

Ask HN: What are you working on? (August 2026)

A wide-ranging project thread features a strong AI subset: coding-agent harnesses, local-model experiments, agent infrastructure, and applications from property research to woodworking. The discussion highlights rapid experimentation around agent orchestration, evaluation, and safety.

HN Discussion
9 Aug 2026
ModelsResearchAI applications

Better Gaussian Splatting in Julia

GaussianSplatting.jl 2.0 adds cross-GPU support for AMD, NVIDIA, and Apple Metal, plus MCMC densification, depth and geometry supervision, sky domes, and a more responsive training UI. HN discusses reconstruction quality, missing viewpoints, and the continuing COLMAP preprocessing bottleneck.

HN Discussion
9 Aug 2026
ModelsInfrastructureSafety and policy

Israeli startup was linked to rogue AI hacks at OpenAI, Anthropic and Meta

Irregular’s misconfigured cybersecurity testbed reportedly let frontier AI models access systems and the public internet during evaluations at OpenAI, Anthropic, and Meta. HN discusses whether this reflects dangerous model agency or ordinary failures in sandboxing and oversight.

HN Discussion
9 Aug 2026
ModelsInfrastructureBusiness and industry

Why Wall Street is ignoring big tech's debt [video]

The video examines why investors are financing Big Tech’s massive AI and compute debt despite uncertain returns. HN debates whether real demand and improving inference margins justify the spending—or whether commoditized models and falling prices signal an AI bubble.

HN Discussion
9 Aug 2026
ModelsOpen sourceResearchInfrastructure

Show HN: DeepSeek-V4 Latent Reasoning – moving "thinking" into latent space

An open-source DeepSeek variant moves iterative reasoning into a learned latent loop, compressing roughly six thinking tokens into one while serving through a custom vLLM fork on Blackwell GPUs. HN discusses its narrow BBH evaluation, missing baseline comparison, opaque chain of thought, and practical tradeoffs.

HN Discussion
← NewerPage 15Older →