source&pool
A daily wire of long-form journalism, video, and discourse — filed, tagged, and laid out flat.
VOL. I·NO. 01
WEDNESDAY, SEPTEMBER 16, 2026
Hacker News3607X 主题热门3491MacRumors78CNBC71YahooFinance649to5Mac59Kotaku44Verge42IGN339to5Google32aihot31Gematsu30NintendoLife30TechCrunch25Engadget24Eurogamer24BusinessInsider23Guardian20CNET15NBC15FoxBusiness14NPR14Fortune13Polygon13SeekingAlpha13bgr12Gizmodo12CBS11Wccftech11Investor'sBusinessDaily10Mashable10TechPowerUp10USAToday10WIRED10PushSquare9CNN8NintendoEverything8Notebookcheck8NewYorkPost8CrudeOilPricesToday8VideoGamesChronicle8ABC7ArsTechnica7Fox7GameInformer7WindowsCentral7BleepingComputer6Deadline5GamesIndustry.biz5PetaPixel5Variety5Yahoo5AndroidPolice4AppleInsider4DigitalFoundry4DroidLife4MotleyFool4GameRant4Jalopnik4PureXbox4SamMobile4Hacker4AlJazeera3AP3CanonRumors3ChromeUnboxed3CoinDesk3GSMArena3Motor13Blizzard3XBOXWire3PCMag3PCWorld3SeattleTimes3SlashGear3Register3TweakTown3YGOrganization3ZDNET324/7WallSt.2Aftermath2AndroidCentral2AwfulAnnouncing2BleedingCool2BuzzFeed2CTech2DualShockers2DW2EventHubs2Futurism2GameDeveloper2Hodinkee2Independent2Lifehacker2MassivelyOverpowered2MyNintendo2Nature2Newser2Newsweek2PaulKrugman2PokémonGOHub2RoadtoVR2RPGSite2Space2Conversation2NextWeb2Tom'sGuide2UploadVR2VideoCardz2WarhammerCommunity2WindowsLatest2YourTango2404Media143rumors1ABC111AboveLaw1ageofempires1AndroidHeadlines1AOL1AVClub1Benzinga1BikeRadar1Billboard1BloodyDisgusting1Borderlands1Bungie1Yahoo!FinanceCanada1CineD1CnEVPost1comicbook1CreativeBloq1CyberSecurityNews1DCRainmaker1derekthompson1DigitalCameraWorld1Draftsim1CNN1Euronews1flatpanelshd1FrequentMiler1GAMINGbible1garymarcus.substack1GearPatrol1GeekWire1GeekyGadgets1Hackaday1HollywoodReporter1InsiderGaming1InterconnectsAI1InterestingEngineering1JapanTimes1KITCO1KrebsonSecurity1KSL1LosAngelesTimes1Lloyd'sList1WPLGLocal101Macworld1Maxroll1Mediaite1MiddleEastEye1MonochromeWatches1MPR1SemiAnalysis1Newsshooter1NoMan'sSky1nylon.com.sg1NYT1OregonLive1PCGamesN1PersonaCentral1Pokemon1politico.eu1PittsburghPost-Gazette1QuantaMagazine1qz1RockPaperShotgun1SammyGuru1ScienceAlert1ScientificAmerican1SouthChinaMorningPost1Semafor1SFGATE1YahooFinanceSingapore1YahooSingapore1SportsIllustrated1SimpleFlying1Sources1supercarblondie1Tedium1TelecomTalk1GameBusiness1TheGamer1Intercept1Times1LongmontTimes-Call1TmoNews1TopGear1TwistedVoxel1YahooFinanceUK1UnHerd1vox1WhatHi-Fi?1WPBF1WRAL1
  1. 001aihotSEP · 12English

    DeepSeek 发布 V4.1-Flash,大幅降低 AI Agent 的 KV cache 内存需求

    DeepSeek released V4.1-Flash, a new AI model that significantly reduces memory requirements for AI agents by shrinking the KV cache to about a quarter of its predecessor's size. The model uses 552 billion parameters and employs techniques like splitting the architecture into encoder and decoder components to halve compute needs for input processing. Performance matches leading models on coding tasks, though weaknesses remain in scientific reasoning and image analysis.

  2. 002Hacker NewsSEP · 12English

    LRU is harder to beat than the KV-cache papers suggest

    A researcher replayed real Claude Code and Mooncake requests through a prefix-cache simulator to test whether alternative eviction policies could beat LRU, but failed across three approaches. The analysis reveals that under capacity pressure, most recomputation stems from tool-calling loops seconds apart rather than idle sessions, and the production LRU baseline proves surprisingly difficult to improve upon.

    By Gauravapiscean
  3. 003Hacker NewsSEP · 11English

    Why Codex burns a weekly limit in a day while the agent waits for the tide

    A token leak in Codex's goal mode causes users to burn through $200 weekly subscription limits in a day. The issue stems from the orchestrator restarting the model every 0.03 seconds while waiting for child agents, forcing it to reread 120–470k tokens per continuation until hitting rate limits. Proposed fixes include using models with native sleep support or implementing pauses between continuations.

    By Ivan Oparin; Alexis Grigoryev; Timur Kackan
  4. 004Hacker NewsSEP · 09English

    LatentMathBench: Investigating Latent Reasoning in Astra

    LatentMathBench is a benchmark designed to test whether large language models like OpenAI's GPT-6 Astra can perform long chains of sequential reasoning in latent space rather than through visible chain-of-thought. The benchmark uses recurrent depth architecture with potential KV-cache sharing, which could allow reasoning to happen opaquely in cached states while producing arbitrary filler text, making the model difficult to monitor.

    By MaartenBaert
  5. 005Hacker NewsSEP · 09English

    Zero-Copy KV-Cache Migration Protocol (81.6ms Latency)

    An open-source protocol for migrating Large Language Model KV-Cache states across datacenters, achieving 81.73ms latency and reducing GPU compute overhead by 95% compared to standard re-computation methods. The zero-copy transport eliminates prompt re-computation during session handoffs, enabling faster Time-To-First-Token latency.

    By DOMINICALI1