Eieye
Eieye
FeedChatReportsAbout
ModelsAgentsCoding toolsOpen sourceResearchInfrastructureAI applicationsSafety and policyBusiness and industry
Filter stories by topic
ModelsAgentsCoding toolsOpen sourceResearchInfrastructureAI applicationsSafety and policyBusiness and industry

Updated 1 Sept, 17:26

13 Aug 2026
ModelsOpen sourceInfrastructure

AI At Home Part 1: A Box Of Scraps

A developer builds a four-GPU home inference server from used AMD hardware, custom cooling, and salvaged parts. HN discusses ROCm support, local model performance, and whether self-hosting is worth the cost versus cloud APIs.

HN Discussion
13 Aug 2026
AgentsOpen sourceInfrastructure

We eliminated 1,400 CVEs in NanoClaw's container images

Echo describes reducing roughly 99% of the reported CVEs in NanoClaw’s container through dependency upgrades, OS package patching, and AI-assisted security backports. HN debates whether the raw CVE count is meaningful and whether agent-runtime security claims address prompt injection.

HN Discussion
13 Aug 2026
ModelsOpen sourceInfrastructureAI applications

I built a 500k-domain search engine for makers in a weekend for $10

An open-source personal search engine uses a small local language model to summarize and categorize 560,000 domains for about $10 in GPU time. The build report and discussion examine crawl steering, model-generated taxonomy problems, local inference costs, and AI-assisted content concerns.

HN Discussion

Built by Will Etheridge

wjeth.comwjeth@pm.me
13 Aug 2026
AgentsCoding toolsOpen source

DeepSeek Harness developer preview

DeepSeek open-sources a developer-preview agent harness built around modular, hot-loadable plugins for models, tools, sessions, and UI. HN discusses its traceable event log, compatibility with other models, and whether the Cordis architecture offers meaningful advantages over existing coding agents.

HN Discussion
12 Aug 2026
ModelsAgentsOpen sourceInfrastructure

Qwen3.8-2.4T

Qwen releases FP8-quantized weights for its 2.4T-parameter Qwen3.8 model, with 95B parameters activated and support for vLLM, SGLang, and other inference stacks. HN discusses its agent and coding benchmarks, upcoming smaller variants, and the enormous hardware required to run it.

HN Discussion
12 Aug 2026
AgentsOpen sourceAI applications

Show HN: OJCP – an open protocol for agent-consumable job data

OJCP proposes an MCP-native standard for AI agents to discover jobs, apply on candidates’ behalf, and track applications using structured schemas, consent, and identity proofs. HN discusses protocol governance, adoption, AI screening, and the changing hiring ecosystem.

HN Discussion
12 Aug 2026
ModelsOpen sourceResearchInfrastructure

Qwen3.8-2.4T

Alibaba has released Qwen3.8, a 2.4T-parameter mixture-of-experts model with 95B active parameters, long-context support, and claimed frontier-level coding and agent performance. HN focuses on its enormous serving requirements, missing vision support, quantization trade-offs, licensing, and implications for open model competition.

HN Discussion
12 Aug 2026
ModelsOpen source

Qwen3.8 Weights Released

Qwen3.8 model weights are available, including a large MoE version and a more accessible 27B variant. HN commenters discuss which versions can run on consumer GPUs and whether a smaller MoE model may follow.

HN Discussion
12 Aug 2026
AgentsCoding toolsOpen sourceAI applications

Hax – a minimalist, terminal-native coding agent written in C

Hax is a minimalist C-based terminal coding agent designed for local models, low resource use, and inspectable tool interactions. HN users discuss its lean design, provider compatibility, model behavior, and hands-on coding results.

HN Discussion
12 Aug 2026
ModelsOpen sourceAI applications

Qwen 3.8-27B goes openweight in 2 days

Qwen is preparing to release the openweight Qwen3.8-27B, a vision-language model aimed at coding, reasoning, and long-horizon agent tasks. HN discusses the withdrawn countdown links, local hardware requirements, and whether 27B models are practical for real-world agents.

HN Discussion
12 Aug 2026
AgentsOpen sourceInfrastructureAI applications

My Agent Setup

A founder details a six-agent AI staff running on a small VPS, coordinated through open-source Buzz and Nostr, with agents handling development, operations, research, marketing, and administration. HN discusses the setup’s security, privacy, cost, reliability, and disappointing early ROI.

HN Discussion
12 Aug 2026
AgentsCoding toolsOpen sourceAI applications

Bb: The IDE that builds itself

bb is a local-first, MIT-licensed development workbench that hosts multiple coding agents and lets them customize the IDE itself through plugins and prompts. HN commenters debate whether it qualifies as an IDE and discuss the benefits of combining agents with highly programmable environments like Emacs.

HN Discussion
12 Aug 2026
ModelsOpen sourceInfrastructure

llama.cpp

The llama.cpp project’s new llama.app experience brings simpler installation and local model serving to its fast, hardware-flexible LLM runtime. HN discusses backend performance, multi-model serving, installation security, and whether it offers advantages over Ollama.

HN Discussion
11 Aug 2026
ModelsAgentsOpen sourceInfrastructure

Nvidia Nemotron 3.5 Lightning and NeMo Switchyard

NVIDIA released the open 30B MoE Nemotron 3.5 Lightning for efficient agentic workloads, alongside NeMo Switchyard for routing tasks across models. HN users examine its local performance, MoE trade-offs, caching challenges, benchmark claims, and rough deployment experience.

HN Discussion
11 Aug 2026
ModelsAgentsOpen sourceInfrastructure

Nvidia Nemotron 3.5 Lightning

NVIDIA released Nemotron 3.5 Lightning, an open 30B/3B-active hybrid Mamba-MoE model with a 1M-token context window and optimized local inference. HN discusses its open training recipe, FP4 performance tradeoffs, comparisons with Qwen, and hands-on coding-agent tests.

HN Discussion
11 Aug 2026
AgentsCoding toolsOpen source

Closing Canario Terminal source code

The Canario terminal is going closed source after its maintainer was overwhelmed by support demands and AI-generated issues and pull requests. HN debates whether read-only publishing could preserve openness while avoiding the new contribution firehose.

HN Discussion
11 Aug 2026
ModelsOpen sourceInfrastructure

No, local models will not win

An argument that local AI models will remain a niche because frontier capability and datacenter batching make cloud inference more efficient. HN commenters challenge the “strongest model always wins” assumption and discuss cost, freedom, privacy, and specialized local use cases.

HN Discussion
10 Aug 2026
AgentsCoding toolsOpen sourceAI applications

Show HN: AI Pulse a fake LED strip beside the macOS Dock that shows agent status

AI Pulse is an open-source macOS indicator that shows whether local AI coding agents are working, waiting for input, finished, or failed. It integrates with Claude Code and pi through a localhost API and displays status in a Dock-adjacent virtual LED strip.

HN Discussion
10 Aug 2026
ModelsAgentsOpen sourceInfrastructure

Show HN: Needle2: 14MB agentic LLM for phones, wearables, smart home and robots

Cactus released Needle 2, a 14MB, 2-bit model that runs in 28MB of RAM for fast local tool calling and structured extraction on inexpensive edge devices. HN users praised its WASM and microcontroller potential but found brittle intent handling and questioned whether its confidence scores reliably detect failures.

HN Discussion
10 Aug 2026
ModelsOpen sourceSafety and policyBusiness and industry

Mark Zuckerberg attacks 'closed' AI rivals as Meta returns to open models

Mark Zuckerberg argues that AI should remain broadly accessible as Meta resumes releasing open-weight models and attacks closed rivals. HN debates whether this expands competition and user control or is mainly a strategic effort to commoditize competitors, alongside concerns about safety and what “open” really means.

HN Discussion
← NewerPage 6Older →