source&pool
A daily wire of long-form journalism, video, and discourse — filed, tagged, and laid out flat.
VOL. I·NO. 01
TUESDAY, SEPTEMBER 29, 2026
X 主题热门3932Hacker News3924CNBC81YahooFinance80aihot77Verge64IGN509to5Mac419to5Google38MacRumors38Engadget37Kotaku35TechCrunch34AndroidAuthority30PushSquare25Eurogamer21Guardian21NintendoLife20Investor'sBusinessDaily19ArsTechnica18Mashable18TechPowerUp17FoxBusiness14Polygon14Wccftech14Gematsu13SeekingAlpha13BusinessInsider12Fortune11Gizmodo11NBC11NPR11bgr10CNN10VideoGamesChronicle10WIRED10GSMArena9USAToday9AlJazeera8AndroidPolice8DroidLife8GameGPU8GameInformer8NewYorkPost8PureXbox8CBS7CNET7MotleyFool7Fox7GamesIndustry.biz7XBOXWire7BleepingComputer6HollywoodReporter6Jalopnik6NintendoEverything6Hacker6Yahoo6ABC5AndroidCentral5AppleInsider5InsiderGaming5Tom'sGuide5VideoCardz5Draftsim4GAMINGbible4NYT4CrudeOilPricesToday4PokeBeach4SamMobile4Space4UploadVR4WarhammerCommunity4WhatHi-Fi?4404Media3Aftermath3Electrek3EventHubs3Futurism3Hackaday3Lifehacker3Nature3Notebookcheck3PCMag3PetaPixel3PlayStationLifeStyle3SouthChinaMorningPost3Yahoo3YahooTech3TechSpot3Conversation3Register3WindowsCentral324/7WallSt.2AndroidHeadlines2Anthropic2Autonocion2AZFamily2BostonGlobe2BuzzFeed2ChromeUnboxed2CoinDesk2Currently2DCRainmaker2Deadline2DigitalFoundry2Euronews2GameRant2GearPatrol2GeekyGadgets2iLovetheUpperWestSide2LosAngelesTimes2Motor12MP1st2mtgrocks2MyNintendo2Phoronix2Pokemon2politico.eu2RoadtoVR2RockPaperShotgun2RPGSite2SeattleTimes2SFGATE2SimpleFlying2SimsCommunity2SlashGear2Slate2Autopian2TimeExtension2TODAY2ynetnews26abcPhiladelphia180Level1WXLV1ageofempires1airlive1AJC1AlaskaBeacon1Apple1AVClub1AviationWeek1Benzinga1BoingBoing1Yahoo!FinanceCanada1YahooLifestyleCanada1CarandDriver1CineD1CnEVPost1comicbookmovie1Skin.ClubCommunity1consequence1CreativeBloq1YahooCreators1DailyKos1DaringFireball1Deseret1Designboom1Dezeen1DigitalCameraWorld1DSOGaming1DiarioAS1Finbold1ForexFactory1FOX13Seattle1franchisetimes1Futurity1GameFile1GameWorldObserver1garymarcus.substack1GeekWire1GosuGamers1Gothamist1Hackster.io1HuffPost1Independent1InsideEVs1InterestingEngineering1investor.costco1Invezz1I/OFund1iPhoneinCanada1KCRA1KOMO1MacObserver1Magic:Gathering1MakeUseOf1Mashed1Minecraft1MLive1MortgageDaily1Motorsport1MyNorthwest1NBCBayArea1NBC5Chicago1NBC7SanDiego1BloombergLaw1Newser1SemiAnalysis1Newsshooter1Newsweek1NintendoWire1Nokiamob1OregonPublicBroadcasting1OregonLive1PageSix1PersonaCentral1PickupTruck+SUVTalk1Psyche1QuantaMagazine1Realtor1RichmondTimes-Dispatch1Richmonder1Road&Track1ScienceDaily1ScreenRant1SeattleRed1SanFranciscoChronicle1YahooFinanceSingapore1GhostHowls1SlowBoring1SoraNews241SpaceNews1statnews1svg1TampaBayTimes1the5krunner1DailyBeast1Intercept1NextWeb1https://tipswatch.com/1TMZ1TopGear1TweakTown1YahooFinanceUK1PCMagUK1Variety1Vulture1WCVB1WFMZ1WHYY1WindowsLatest1WrestlingInc.195.5WSB1YGOrganization1
  1. 021Hacker NewsSEP · 28English

    An Internet for the KV Cache: Rethinking Classical Infrastructure Boundaries

    A research paper proposes decoupling compute and KV Cache storage across cloud infrastructure to optimize LLM inference at scale. The approach treats KV Cache management as a content-distribution system, enabling adaptive decisions based on network bandwidth, latency, and pricing to minimize latency and cost.

    By Ray; Siddhant; Feamster; Nick; Jiang; Junchen
  2. 022X 主题热门SEP · 28English

    "GPU demand" · X 热门 · 2026-09-28 18:02 UTC

    Nvidia authorized a $150B stock buyback, bringing total authorization to $235B, signaling management confidence in sustained AI demand and the durability of the GPU cycle. The company is expanding beyond GPUs into networking, security, orchestration, and infrastructure for agentic AI, allowing it to monetize more of each AI rack while funding aggressive infrastructure development and shareholder returns simultaneously.

  3. 023Hacker NewsSEP · 28English

    Model Release: Naive-N0.5-Flash

    NaiveAI released Naive-N0.5-Flash, a 309B MoE model optimized for coding and AI R&D, built using AI-centered development where models assist in their own design and training. The model features a 1M context window with hybrid attention architecture and achieves up to 2,000 tokens/s inference speed through the NaiveRT system. Weights and inference code are open-sourced under MIT license with API pricing available.

    By NaiveAI Team
  4. 024Hacker NewsSEP · 28English

    Investigating Jev's architecture and how to scale it

    Jev is a System One Model by TypeSafe AI designed for fast structured decision-making without autoregressive generation. It accepts state context, questions (choice, score, or binary), and criteria to produce probability distributions, level classifications, or binary outputs, enabling flexible multi-question requests within context limits.

    By Sriram Govindan
  5. 025Hacker NewsSEP · 28English

    Gevva0 – a Jev like decision engine on Gemma 26B via direct logit scoring

    Gevva0 is a local decision engine built on Gemma 26B that replaces slow autoregressive JSON generation with direct logit scoring and calibration techniques, delivering deterministic classifications with audit trails in 15–45ms. It addresses four failure modes in LLM classification: logit poisoning, positional bias, poor calibration, and latency, using cyclic debiasing and Platt temperature scaling.

    By Solvingsteve
  6. 026Hacker NewsSEP · 28English

    Show HN: Jeva.cpp – a llama.cpp fork with JEV-compatible API for all LLMs

    Jeva.cpp is a llama.cpp fork that adds a JEV-compatible decision API to llama-server, enabling Choice, Score and Noul evaluations from model logits while maintaining standard autoregressive generation across all models and platforms supported by llama.cpp.

    By PragmaTwice
  7. 027Hacker NewsSEP · 28English

    I got 2.2x more tokens per second from llama.cpp on Intel Arc

    A developer benchmarked llama.cpp on an Intel Arc-equipped laptop and found that disabling CPU offload for mixture-of-experts layers achieved 2.2–2.3x faster performance, though it requires careful memory management. Key optimizations included using speculative decoding with n=2 and increasing batch size to 2048 for long prompts, while thread count and CPU governor had negligible impact.

    By Luigi
  8. 028X 主题热门SEP · 28English

    HBM demand · X 热门 · 2026-09-28 09:51 UTC

    A discussion on HBM (High Bandwidth Memory) demand for AI agents, analyzing how personal agent workloads will evolve and drive increased CPU, DRAM, and NAND requirements as users delegate more complex tasks over time, with implications for data centers and GPU deployments.

  9. 029X 主题热门SEP · 28English

    data center revenue · X 热门 · 2026-09-28 09:50 UTC

    Truist projects AI cloud revenue will grow 7x from $310B in 2025 to $2.1T by 2030, driven primarily by inference workloads rather than training. AI cloud providers like Nebius and CoreWeave are expected to capture roughly 30% of the market by 2030, with AI data center capacity needing to expand from 24 GW to 100 GW, requiring significant power infrastructure investment.

  10. 030X 主题热门SEP · 28English

    DeFi · X 热门 · 2026-09-28 08:43 UTC

    DeFi discussions on X cover Base's all-time high activity, cross-chain agent infrastructure via AACP, strkBTC enabling Bitcoin utility on Starknet, and Bittensor's TAO token potentially gaining value from AI inference cost reductions demonstrated by SOMA subnet 114.

  11. 031Hacker NewsSEP · 28English

    Foundations of Large Language Models

    A foundational textbook on large language models covering pre-training, generative models, prompting, alignment, inference, and reasoning. Designed for computer science students, professionals, and NLP practitioners seeking to understand core concepts in the field.

    By Xiao; Tong; Zhu; Jingbo
  12. 032Hacker NewsSEP · 28English

    Build a Real‑Time Telemetry Pipeline for SpaceX Starship Launches

    A technical guide for building a cloud-native streaming pipeline to ingest, process, and monetize telemetry from SpaceX's Starship launch on September 28, 2026, handling up to 5 Gbps of data with sub-second latency for AI safety checks and real-time customer billing.

    By Dheeraj Ramasahayam
  13. 033X 主题热门SEP · 27English

    data center revenue · X 热门 · 2026-09-27 22:56 UTC

    Social media discussion on AI data center revenue growth, with Viavi highlighting 40% revenue increases in optical testing equipment, while analysts project AI cloud revenue rising from $310B in 2025 to $2.1T by 2030, driven primarily by inference workloads requiring massive infrastructure expansion and power generation capacity.

  14. 034X 主题热门SEP · 27English

    AI infrastructure stocks · X 热门 · 2026-09-27 22:23 UTC

    X posts discuss AI infrastructure investment opportunities, focusing on projections that AI cloud revenue could grow from $310 billion in 2025 to $2.1 trillion by 2030, with inference becoming the primary driver of demand. Analysts highlight beneficiary stocks across cloud providers, memory, networking, power generation, and data center infrastructure sectors, while noting emerging opportunities in AI agent platforms and specialized compute providers.

  15. 035Hacker NewsSEP · 27English

    How Jev works: calibrated decision models

    Jev is a decision model from TypeSafe AI that selects among predefined options by scoring their likelihood as text completions, then normalizing scores into probabilities—without generating text. Experiments on Qwen2.5 models show scoring is 7–54× faster than text generation, with comparable accuracy, though performance degrades as option counts increase. Small models can be fine-tuned efficiently to improve calibration and accuracy on specific tasks.

    By Victor Dibia
  16. 036Hacker NewsSEP · 27English

    Is AGENTS.md useful? What the research found

    Research from five 2026 studies on AGENTS.md context files for coding agents shows mixed results: hand-written files provide modest benefits (around 2-4 percentage points) for specifying non-standard practices, while generated files slightly hurt performance, with all approaches increasing inference costs by roughly 20 percent. File structure and organization matter less than which rules are presented and when, and context files outperform skills only for knowledge models lack and need on most tasks.

    By shoustak
  17. 037Hacker NewsSEP · 27English

    Ask HN: How are you getting inference for personal projects?

    A question about how developers are obtaining AI inference for personal projects as major providers reduce included usage in monthly subscriptions. The asker explores options including API subscriptions, hosted open-weight models, and local inference.

    By variety8675
  18. 038Hacker NewsSEP · 27English

    Jev in your local browser in Go with WebGPU using Jeyzma

    Jeyzma is a browser-based System One machine learning model written in Go and compiled with TinyGo, running inference locally via WebAssembly with WebGPU GPU acceleration. It answers typed questions by providing probability distributions for options, with no server or data transmission required.

    By deadprogram
  19. 039Hacker NewsSEP · 27English

    How do you run System One decision models locally?

    System One decision model runtimes load and serve model weights locally on your hardware. As of September 2026, Ollaya is the fastest path, with alternatives like laya-mlx for Apple silicon and llama.cpp for GGUF models. Runtime choice affects model output confidence scores, so version pinning matters for threshold-based decisions.

    By Sergei Gordeichuk
  20. 040Hacker NewsSEP · 27English

    Show HN: Peekaboolean – image and jev-like typed questions in typed anwers out

    Peekaboolean is a vision-language model that answers typed questions about images using choice, score, and noun formats similar to Jev. The system encodes images once on a server and scores options as short suffixes, achieving ~400ms p95 latency on M1 Pro and ~60ms on desktop GPU.

    By bykof