source&pool
A daily wire of long-form journalism, video, and discourse — filed, tagged, and laid out flat.
VOL. I·NO. 01
SUNDAY, OCTOBER 11, 2026
Hacker News1670X 主题热门1501ChainCatcher111CNBC75PANews67Coinness53Channel News Asia49SouthChinaMorningPost35Crypto.news30YahooFinance30CoinDesk21Verge21Kotaku20Cointelegraph19aihot17Variety17IGN169to5Mac15MacRumors15TechCrunch129to5Google11AMBCrypto11Decrypt11DW11TechPowerUp11Eurogamer10Guardian10FoxBusiness9NintendoLife9Wccftech9BBC World8BusinessInsider7VideoCardz7BitcoinMagazine6Engadget6NBC6Polygon6WarhammerCommunity6ArsTechnica5CNET5Gematsu5Gizmodo5Investor'sBusinessDaily5XBOXWire5Register5bgr4CBS4CNN4Futurism4GSMArena4Mashable4Notebookcheck4NYT4PushSquare4USAToday4AlJazeera3AppleInsider3MotleyFool3Fortune3Fox3PokémonGOHub3SeekingAlpha3Hacker3VideoGamesChronicle3BleepingComputer2DigitalFoundry2DroidLife2DSOGaming2Euronews2KSL2Lifehacker2MyNintendo2Nature2Newser2NPR2SeattleTimes2Yahoo2WindowsCentral2WIRED2Yahoo224/7WallSt.16abcPhiladelphia1ABC7NewYork1ABC1AlineaInsightnewsletter1AndroidCentral1AndroidPolice1NikkeiAsia1Bank of England1Barron's1Beebom1BloodyDisgusting1BostonGlobe1CFTC1ChromeUnboxed1Cleveland1CreativeBloq1EventHubs1Federal Reserve1FTC1GameDeveloper1GameFile1GameRant1GeekWire1HotHardware1HouseDigest1Independent1InsiderGaming1InvenGlobal1KTLO1KUTV1MP1st1NBC5Chicago1NBCSports1MicrosoftSource1NintendoWire1CrudeOilPricesToday1OMG!Ubuntu1PaulKrugman1PCGamer1PennLive1Phoronix1Pocket-lint1Pokemon1PureXbox1OutlookRespawn1RetractionWatch1Road&Track1RockPaperShotgun1RPGSite1ScienceAlert1SEC1SFGATE1YahooFinanceSingapore1SpaceNews1Syracuse1TimesSquareChronicles1YahooTech1Hill1Outerhaven1Times1Tom'sGuide1TopGear1TweakTown1YahooUK1OutsideMagazine1VGChartz1EdZitron'sWhere'sYourEdAt1WolfStreet1YankoDesign1ZDNET1
  1. 001Hacker NewsOCT · 11English

    The BM25 Weighting Scheme

    BM25 is a probabilistic weighting scheme used by Xapian that combines previous models (BM11 and BM15) with a scaling factor to improve term weighting in information retrieval. Recent TREC tests have shown it to be the best known probabilistic weighting scheme, with default parameters of k1=1, k2=0, k3=1, and b=0.5 that can be tuned for specific collections and query types.

    By ankitg12
  2. 002Hacker NewsOCT · 10English

    Using KV cache as embeddings

    BreadBowl-Embed proposes a new retrieval representation that sits between single-vector and token-level approaches, using 16 fixed slots per passage with dual vectors for routing and value reading. This addresses the inefficiency of two-stage RAG pipelines where documents are read twice—first for retrieval via bi-encoder, then for reranking via cross-encoder—reducing computational waste especially for agents that issue multiple queries.

    By Ming Xu
  3. 003Hacker NewsOCT · 09English

    Efficient Agents

    A comprehensive survey examining efficiency in large language model-based agents, focusing on three core components: memory, tool learning, and planning. The paper reviews approaches for reducing costs such as latency and tokens through techniques like context compression, retrieval, budgeted tool use, and hierarchical planning, while evaluating efficiency through Pareto frontier analysis between effectiveness and cost.

    By FIRST AUTHOR NAME; SECOND AUTHOR NAME; FIRST AUTHOR LAST; FIRST AUTHOR FIRST; SECOND AUTHOR LAST; SECOND AUTHOR FIRST
  4. 004Hacker NewsOCT · 09English

    Stop Looking for the Best Way to Retrieve Context for AI Agents. Build a Router

    Zenith argues that AI agents need routing-based retrieval systems tailored to different question types rather than a single best retrieval method. Vector search alone fails for questions requiring document versioning, relationship paths, or current state; the solution is hybrid retrieval combining vector and BM25 search, with GraphRAG only when relationship queries genuinely outperform tuned baselines.

    By Manveer Chawla
  5. 005Hacker NewsOCT · 09English

    Long-Term Memory for AI:50M-Token Window,Is Faster,Cheaper Than Recompute

    Researchers demonstrate a memory layer that extends AI language models to handle 50-million-token contexts by storing and retrieving key-value states from encrypted disk storage, achieving 2.8-4.3x faster loading and 8.8-12.3x lower GPU energy use compared to recomputation. Testing on Gemma models shows accurate retrieval of facts from millions of tokens earlier with no hallucinations.

    By Schelpe; Sietse
  6. 006Hacker NewsOCT · 08English

    Embedding Gemma2 Use Cases

    Med Karim Bchini ports jeffhub.ai use cases to Google's EmbeddingGemma-2 embedding model, replacing a 0.8B decider with dense embeddings and cosine similarity scoring. The Node.js implementation runs CPU-only using quantized ONNX weights (~314 MB), achieving ~170 ms latency per embedding and 100% accuracy across classification and retrieval tasks on small curated datasets.

    By karimtn