source&pool
A daily wire of long-form journalism, video, and discourse — filed, tagged, and laid out flat.
VOL. I·NO. 01
THURSDAY, SEPTEMBER 17, 2026
Hacker News3925X 主题热门3797CNBC84MacRumors71YahooFinance619to5Mac58Kotaku45Verge41IGN349to5Google33aihot33Gematsu31NintendoLife31BusinessInsider27Engadget26Eurogamer26TechCrunch26Guardian18Polygon17NBC16CNET14NPR14Fortune13FoxBusiness13USAToday13Wccftech13bgr12Gizmodo12Mashable12PushSquare12SeekingAlpha12Notebookcheck10WIRED10TechPowerUp9ABC8AppleInsider8CNN8Fox8GameInformer8Investor'sBusinessDaily8NewYorkPost8VideoGamesChronicle8WindowsCentral8CBS7ArsTechnica6BleepingComputer6XBOXWire6NintendoEverything6CrudeOilPricesToday6Variety6AndroidPolice5GamesIndustry.biz5PetaPixel5PureXbox5SamMobile5AlJazeera4AndroidAuthority4CoinDesk4DigitalFoundry4GameRant4GSMArena4SlashGear4Conversation4Register4WarhammerCommunity4Yahoo4CTech3ChromeUnboxed3Deadline3DW3Jalopnik3Lifehacker3Motor13Blizzard3PCMag3PCWorld3Pokemon3RockPaperShotgun3RPGSite3SouthChinaMorningPost3SeattleTimes3Space3Hacker3TweakTown3VideoCardz3WindowsLatest3ZDNET3404Media280Level2Aftermath2AndroidCentral2AOL2AwfulAnnouncing2BleedingCool2BuzzFeed2CanonRumors2CyberSecurityNews2DroidLife2DualShockers2Euronews2EventHubs2MotleyFool2FratelloWatches2Futurism2GameDeveloper2GearPatrol2Hodinkee2KITCO2LosAngelesTimes2MassivelyOverpowered2Maxroll2MP1st2MyNintendo2Nature2Newser2PaulKrugman2PokémonGOHub2RoadtoVR2SFGATE2Intercept2NextWeb2Tom'sGuide2UploadVR2YourTango2ABC111AboveLaw1ageofempires1AVClub1Benzinga1BikeRadar1Billboard1BloodyDisgusting1Borderlands1Boston1Bungie1Yahoo!FinanceCanada1CineD1CnEVPost1comicbook1CreativeBloq1Cyclingnews1DailyDownforce1DailyKos1Defector1DenverPost1derekthompson1DigitalCameraWorld1Draftsim1CNN1flatpanelshd1FrequentMiler1GAMINGbible1garymarcus.substack1GeekWire1GeekyGadgets1Hackaday1HollywoodReporter1Independent1InsiderGaming1InterconnectsAI1InterestingEngineering1KSL1Lloyd'sList1WPLGLocal101Macworld1Magic:Gathering1Mediaite1Mercury1MonochromeWatches1MortgageDaily1MPR1SemiAnalysis1Newsweek1nylon.com.sg1NYT1OregonLive1PageSix1PCGamesN1politico.eu1PittsburghPost-Gazette1QuantaMagazine1qz1SammyGuru1CultureMapSanAntonio1ScienceAlert1ScientificAmerican1Semafor1YahooFinanceSingapore1YahooSingapore1SportsIllustrated1SimpleFlying1supercarblondie1YahooTech1Tedium1TelecomTalk1GameBusiness1TheGamer1TimeExtension1LongmontTimes-Call1TmoNews1TopGear1TwistedVoxel1YahooFinanceUK1UnHerd1VisualCapitalist1WhatHi-Fi?1WOWT1WPBF1WRAL1WSB-TV1YGOrganization1
  1. 001Hacker NewsSEP · 17English

    VC-Attention: Faster Low-Bit Attention Without Retraining

    Nunchux introduces VC-Attention, a training-free low-bit attention method that accelerates video generation by 1.6× on NVIDIA B200 compared to BF16 FlashAttention-4. The technique combines V-Smooth for reducing quantization error and ExpCast-FP8 for eliminating the softmax bottleneck, achieving higher fidelity than existing methods like SageAttention2.

    By Nunchux AI
  2. 002Hacker NewsSEP · 16English

    SGLang and Miles Add Day-0 Support for DeepSeek-v4.1

    SGLang and Miles add day-0 support for DeepSeek-V4.1, a model featuring low-ratio compression, sliding-window attention, manifold hyper-connections, and Engram memory for efficient serving. The implementation includes cross-layer sharing, sparse retrieval mechanisms, and host-memory placement optimizations that increase KV cache capacity by 36% while maintaining comparable throughput.

    By SGLang; Miles Teams
  3. 003Hacker NewsSEP · 16English

    Effectively Does a Model Use Its Memory (2025)

    This paper introduces Effective State-Size (ESS), a metric measuring how well sequence models utilize their memory by analyzing the rank of input-dependent transformation matrices. ESS reveals that models with high memory utilization are harder to distill, and that effective ESS modulation correlates with better performance on recall-intensive tasks.

    By Liquid AI
  4. 004Hacker NewsSEP · 16English

    Singapore launched a novel five-year national campaign to nurture reading habits

    Singapore launched a five-year national campaign to encourage reading habits by offering small monetary rewards to citizens who log reading sessions on a government website. The initiative uses gamification and aims to combat declining attention spans and the rise of shallow online content consumption, though it has sparked debate about whether financial incentives can truly foster a genuine love of reading.

    By Kathleen Magramo
  5. 005Hacker NewsSEP · 15English

    LLM Speedrun: Architecture

    Article explaining LLM architecture fundamentals, focusing on the transformer model and attention mechanism. Covers how transformers parallelize computation compared to RNNs, and how attention allows tokens to dynamically reference all previous context. Includes code examples and notation for understanding embeddings, queries, keys, and values.

    By bucket2015
  6. 006Hacker NewsSEP · 15English

    Against Waldenponding (2018)

    A 2018 essay critiques Waldenponding—the philosophy of retreating from technology to reclaim attention—arguing it's counterproductive both personally and collectively. The author contends that staying plugged into information flows and managing attention actively is preferable to unplugging, and that retreat exploits Fear Of Being Ordinary (FOBO) just as social media exploits Fear Of Missing Out (FOMO).

    By Venkatesh Rao
  7. 007Hacker NewsSEP · 15English

    The harms of short-form video on the brain are starting to show

    Research increasingly shows that short-form video consumption is associated with cognitive harms, particularly affecting attention, impulse control, and memory, even as social media companies dispute the scientific evidence. A meta-analysis of 70 studies found stronger associations between short-form video use and negative cognitive changes than with mental health issues, with compulsive use predicting the worst outcomes.

    By Caty Enders
  8. 008Hacker NewsSEP · 14English

    I'm not addicted to the internet or my smartphone. I'm addicted to information

    The author reflects on what they initially perceived as internet and smartphone addiction, ultimately concluding they are actually addicted to information discovery. They describe how algorithmic recommendations and short-form content fail to deliver the meaningful, life-changing information they once found online, and express nostalgia for the exploratory internet of the past.

    By jyhrow
  9. 009Hacker NewsSEP · 14English

    Understanding FlashAttention Pt 1: Personal Notes

    A technical handbook explaining FlashAttention, an optimization technique that accelerates transformer attention mechanisms through tiling, online softmax, and recomputation without approximating the mathematical function. The key insight is that wall-clock speed depends on GPU memory traffic rather than FLOP count alone, achieved by reducing expensive reads and writes to high-bandwidth memory.

    By Chizkidd
  10. 010Hacker NewsSEP · 14English

    OpenArch – PyTorch implementations of modern LLM architectures

    OpenArch is a PyTorch repository containing hand-written implementations of modern LLM architectures designed for educational clarity rather than production performance. Each model is implemented from scratch in a single readable file, making architectural choices like attention types, normalization methods, and positional encodings explicit and easy to compare across 72 different architectures.

    By Anuj
  11. 011Hacker NewsSEP · 13English

    MOBA: Mixture of Block Attention for Long-Context LLMs

    MoBA (Mixture of Block Attention) is a novel attention mechanism for long-context LLMs that applies Mixture of Experts principles to reduce computational complexity while allowing models to autonomously determine attention patterns. The approach enables seamless transitions between full and sparse attention and has been deployed in Kimi's long-context system.

    By Lu; Enzhe; Jiang; Zhejun; Liu; Jingyuan; Du; Yulun; Tao; Hong; Chao; Shaowei; He; Weiran; Yuan; Enming; Wang; Yuzhi; Huang; Zhiqi; Xu; Suting; Xinran; Lai; Guokun; Chen; Yanru; Zheng; Huabin; Junjie; Jianlin; Wu; Yuxin; Zhang; Neo Y; Yang; Zhilin; Zhou; Xinyu; Mingxing; Qiu; Jiezhong
  12. 012Hacker NewsSEP · 13English

    Recurrent Looped Transformer

    Recurrent Looped Transformer (RLT) combines a causal encoder with a recurrent decoder that grows temporal depth with each token, traversing tL_D decoder blocks after t tokens while maintaining fixed per-token computation. The architecture integrates global key–value memory, layerwise sliding-window attention caches, and continuous latent computation across prompts and responses, with co-design considerations for hardware efficiency and reinforcement learning scaling.

    By Yifan Zhang
  13. 013Hacker NewsSEP · 13English

    Neural Turing Machines (2014)

    Neural Turing Machines couple neural networks with external memory resources accessed via attention mechanisms, creating a differentiable system analogous to traditional computing architectures. The approach enables end-to-end training with gradient descent and can learn algorithms like copying, sorting, and associative recall from examples.

    By Graves; Alex; Wayne; Greg; Danihelka; Ivo
  14. 014Hacker NewsSEP · 13English

    Recurrent Looped Transformer

    Recurrent Looped Transformer (RLT) combines a causal encoder with a recurrent decoder that grows temporal depth with each token, traversing tL_D decoder blocks after t tokens while maintaining fixed per-token computation. The architecture integrates global encoder memory with layerwise sliding-window attention caches in the decoder, enabling co-design with hardware and RL algorithms through shared state transitions across pretraining, fine-tuning, and sampling.

    By Yifan Zhang
  15. 015Hacker NewsSEP · 10English

    Looped Transformers as Programmable Computers (2023)

    Researchers demonstrate that transformer networks can function as universal computers by programming specific weights and looping the architecture. Using input sequences as instructions and memory, they show that shallow transformers can emulate computing blocks like branches and function calls, enabling execution of algorithms including calculators, linear algebra operations, and backpropagation-based learning.

    By Giannou; Angeliki; Rajput; Shashank; Sohn; Jy-yong; Lee; Kangwook; Jason D; Papailiopoulos; Dimitris