source&pool
A daily wire of long-form journalism, video, and discourse — filed, tagged, and laid out flat.
VOL. I·NO. 01
SUNDAY, SEPTEMBER 27, 2026
X 主题热门3532Hacker News3508CNBC66YahooFinance58aihot56Verge52IGN459to5Mac41MacRumors37Kotaku30Engadget289to5Google27TechCrunch24AndroidAuthority21Eurogamer21NintendoLife20Guardian19PushSquare18ArsTechnica16Wccftech15FoxBusiness14TechPowerUp14Investor'sBusinessDaily13Polygon13BusinessInsider12USAToday12Fortune11Gematsu11Mashable11CBS10Gizmodo10AndroidPolice9bgr9CNN9NPR9VideoGamesChronicle9CNET8MotleyFool8NBC8NewYorkPost8PureXbox8SeekingAlpha8BleepingComputer7Fox7GSMArena7Notebookcheck7Tom'sGuide7VideoCardz7ABC6NintendoEverything6WarhammerCommunity6AndroidCentral5AppleInsider5CoinDesk5GamesIndustry.biz5XBOXWire5PokeBeach5TechSpot5WIRED5AlJazeera4DroidLife4GameGPU4GameInformer4GAMINGbible4HollywoodReporter4InsiderGaming4PetaPixel4PlayStationLifeStyle4RockPaperShotgun4SlashGear4Conversation4Hacker4Yahoo4Aftermath3ChromeUnboxed3Draftsim3Electrek3EventHubs3Futurism3Hackaday3Jalopnik3Lifehacker3Motor13NYT3SamMobile3SouthChinaMorningPost3UploadVR3WhatHi-Fi?3404Media26abcPhiladelphia2AndroidHeadlines2AZFamily2Benzinga2DCRainmaker2Deadline2GameRant2GearPatrol2MP1st2MyNintendo2Nature2CrudeOilPricesToday2PCMag2Pokemon2RoadtoVR2RPGSite2SFGATE2SimsCommunity2Slate2DailyBeast2Register2TimeExtension2TODAY2Variety2WindowsCentral224/7WallSt.180Level1ABC7LosAngeles1ageofempires1Anthropic1Apple1ArizonaSports1Autonocion1AVClub1BellofLostSouls1BostonGlobe1BusinessTimes1BuzzFeed1Yahoo!FinanceCanada1YahooLifestyleCanada1CarandDriver1CarBuzz1CineD1ClaimDepot1Skin.ClubCommunity1consequence1CreativeBloq1YahooCreators1Currently1DailyKos1DaringFireball1DarkHorizons1Decrypt1Deseret1Designboom1Dezeen1DigitalCameraWorld1DirtonDirt1DSOGaming1ForexFactory1FOX13Seattle1franchisetimes1Futurity1GameWorldObserver1GeekWire1GeekyGadgets1Global1GosuGamers1Gothamist1Hackster.io1HouseDigest1HoustonChronicle1iLovetheUpperWestSide1InsideEVs1InterestingEngineering1investor.costco1Invezz1I/OFund1iPhoneinCanada1KCRA1MacObserver1Magic:Gathering1MakeUseOf1Mashed1MLive1MortgageDaily1Motorsport1mtgrocks1NBCBayArea1NBC5Chicago1BloombergLaw1SemiAnalysis1Newsweek1NintendoWire1OregonPublicBroadcasting1OregonLive1PersonaCentral1Phoronix1PickupTruck+SUVTalk1politico.eu1Psyche1qz1Realtor1Richmonder1Road&Track1ScienceDaily1SeattleRed1SeattleTimes1SanFranciscoChronicle1YahooFinanceSingapore1Yahoo1SimpleFlying1GhostHowls1SlippedDisc1SlowBoring1SoraNews241SpaceNews1statnews1svg1TampaBayTimes1YahooTech1the5krunner1DailyMeal1Drive1Hindu1Intercept1TimesofIndia1TMZ1TopGear1TweakTown1YahooFinanceUK1PCMagUK1Vulture1WCVB1WFMZ1WHYY1WKYT1YGOrganization1
  1. 001Hacker NewsSEP · 27English

    Eikos - OSS Jev-like model

    Eikos is an open-source family of typed-decision models (4B and 27B parameters) released under MIT license for structured question-answering in global finance and trade. Each model returns calibrated probabilities for decision options in a single forward pass, with support for multiple GPU formats and Apple Silicon, plus complete training pipelines and evaluation tools.

    By Caiovicentino
  2. 002Hacker NewsSEP · 27English

    Language Model "Shape"

    Alex Zhang discusses how language model input/output shapes have remained static since ChatGPT, with harnesses designed around autoregressive models rather than vice versa. He argues that alternative model architectures with constrained output spaces—like Jev, which outputs values in [0,1]—could enable more efficient solutions for specific use cases and agent designs.

    By Alex L Zhang
  3. 003Hacker NewsSEP · 26English

    Show HN: Fly Language Model

    Fly Language Model (FLM) is a language model trained from scratch on a connectome-derived recurrent network inspired by fruit-fly brain wiring, allowing inspection of neural states during inference. The project combines anatomical wiring constraints with learned dynamics, tested on WikiText-2 and BabyLM datasets to isolate the computational value of the fly-derived architecture versus conventional baselines.

    By kuberwastaken
  4. 004Hacker NewsSEP · 26English

    Making a small language model behave like a Mythos

    Expensive large language models like Claude Mythos outperform cheaper alternatives on complex, long-form tasks and fact retention, but smaller models like Claude Sonnet handle routine work well. Simple prompting techniques—being specific, providing facts, admitting uncertainty, showing examples, and breaking tasks into steps—can significantly narrow the performance gap.

    By Pooja Phillips
  5. 005Hacker NewsSEP · 26English

    Claude Opus 5.5 Should Raise Your Ambitions

    Anthropic released Claude Opus 5.5, a new model performing at Fable 5.1 level while costing 40% less than Opus 5. The model features improved communication, agentic coding capabilities, and zero data retention options, with pricing at $4/$20 and faster speed options available.

    By Zvi Mowshowitz
  6. 006Hacker NewsSEP · 26English

    Can a language model run in Linux eBPF?

    A research project demonstrates running Qwen3-0.6B language model inference in Linux eBPF kernel space, executing 28 decoder layers with token generation in verified BPF programs while using fixed-point arithmetic and BPF maps for state management. The prototype achieves working forward passes at ~1.2 seconds per token, split between C code for tokenization and weight loading and BPF programs for matrix operations, attention, and token selection.

    By Littlefisher
  7. 007Hacker NewsSEP · 25English

    Synthetic Hospital: Physician-Validated Longitudinal EHR Benchmark

    Synthetic Hospital is an open, synthetic longitudinal EHR benchmark built from public medical-education material with 1,268 patients and 5,602 encounters, designed to overcome privacy barriers while providing verifiable ground truth grounded in standard medical ontologies. Physician reviewers distinguished synthetic records from real charts at near-chance rates, and frontier language models achieve at best a severity-weighted F1 of 0.73 on patient problem list reconstruction, revealing significant gaps in clinical AI performance.

    By Park; Christine; Chen; Valerie; Dettmers; Tim
  8. 008Hacker NewsSEP · 25English

    Your Language Model Is Already a Decision Model

    Language models can function as decision models by using next-token probabilities to select actions from candidate options without requiring decision-specific training. Qwen3.5-9B, tested against Jev 1.13.0, achieved competitive accuracy on multiple benchmarks including WebPRM and DeepSWE tasks, with the approach using rotated letter scores and log probability aggregation across candidate positions.

    By Ntlm
  9. 009Hacker NewsSEP · 25English

    Extracting and Characterizing Hidden Chain-of-Thought in Frontier Models

    Researchers developed a method to extract hidden chain-of-thought reasoning from frontier language models like GPT-6 Astra by using a custom API tool, finding that externalized reasoning matches native performance and reveals systematic differences in how models structure intermediate reasoning across mathematics, science, and code tasks.

    By Luo; Xiaoyu; Ren; Tao; Yu; Wenrui; Li; Qiongxiu; Bjerva; Johannes
  10. 010Hacker NewsSEP · 25English

    Scaling Laws for Neural Language Models (first "scaling laws" paper from 2020)

    This paper establishes empirical scaling laws showing that language model loss follows power-law relationships with model size, dataset size, and compute, spanning over seven orders of magnitude. The research demonstrates that larger models are more sample-efficient and that optimal training involves large models on modest data, stopping before convergence.

    By Kaplan; Jared; McCandlish; Sam; Henighan; Brown; Tom B; Chess; Benjamin; Child; Rewon; Gray; Scott; Radford; Alec; Wu; Jeffrey; Amodei; Dario
  11. 011Hacker NewsSEP · 24English

    Show HN: MEF LLM Studio

    MEF LLM Studio is an educational Windows application designed for beginners to understand how language models work through interactive explanations and hands-on experiments. Users can explore tokenization, training dynamics, model comparison, and text generation without writing code.

    By walti1972
  12. 012Hacker NewsSEP · 24Chinese

    JEV-Star: Low-Cost StarCraft II Control with Language-Model Planning

    JEV-Star is a system that combines language-model planning with StarCraft II control, enabling both macro-level game strategy and micro-level unit management. The framework integrates the Astra planner with JEV action selection across multiple game configurations, achieving wins on hard micromanagement tasks and full-game macro control with realtime performance.

    By Sc
  13. 013Hacker NewsSEP · 24English

    Memory Control Signals Emerge Before Action in Long Horizon Agents

    Researchers discovered that language model agents encode signals for memory management (compression and recall) in their hidden states before taking actions, indicating the models already represent when these operations are needed. They propose PaMER, a framework combining state-guided compression with evidence retrieval that reduces context consumption while maintaining task performance on long-horizon agent benchmarks.

    By Wang; Mingxuan; Yao; Guorun; Luo; Fei; Yinglong; Ning; Chao; Bo; Chen; Hongyue; Ma; Yanbiao; Han; Jungong
  14. 014Hacker NewsSEP · 23English

    LatentPort: Cross-model recurrent state transfer without prefix replay

    LatentPort demonstrates cross-model transfer of recurrent inference state from a 4B to 9B Qwen language model without replaying the source context, using hybrid-state handoff combining translated attention KV cache with Gated DeltaNet persistent-state components. The approach achieves near-native performance with only a 0.076 nats/token excess loss on continuation tasks.

    By Villani; Simon P
  15. 015Hacker NewsSEP · 23English

    HySparse2: Hybrid Sparse Attention with Two-Level KV Sharing

    HySparse2 is a hybrid sparse attention architecture designed for long-context language models that improves efficiency through two-level KV sharing between self-decoder and cross-decoder components. It replaces block-level sparsity with token-level sparsity and enables prefill computation to exit early, reducing computational cost and KV-cache storage while maintaining performance on long-context retrieval and multi-turn agent tasks.

    By Wei; Jianyu; Gao; Yizhao; Zhang; Qihao; Shimao; Tang; Zhengju; Cheng; Yu; Zhou; Shengjie; Jiang; Zihan; Song; Yifan; Hailin; Zhao; Liang; Yang; Bo; Wang; Gang; Cao; Shijie; Luo; Fuli
  16. 016Hacker NewsSEP · 23English

    Nemotron-H: A Family of Accurate, Efficient Hybrid Mamba-Transformer Models

    NVIDIA introduces Nemotron-H, a family of hybrid Mamba-Transformer language models (8B to 56B parameters) designed for efficient inference while maintaining competitive accuracy. The models achieve up to 3x faster inference than pure Transformers, with the 56B variant trained on 20 trillion tokens in FP8 precision and capable of supporting ~1-million-token context windows.

    By Bluestein
  17. 017Hacker NewsSEP · 23English

    Jev in 25 Lines of Python

    A tutorial demonstrates implementing Jev, a decision classification model, in 25 lines of Python using the Qwen3 language model to classify email inputs into categories like legitimate, spam, or phishing by extracting and normalizing token logits into probabilities.

    By Duarte O Carmo
  18. 018aihotSEP · 22English

    Claude Opus 5.5 登顶 Artificial Analysis Intelligence Index,并降价 20%

    Claude Opus 5.5 achieved the top ranking on the Artificial Analysis Intelligence Index and received a 20% price reduction. The model matches GPT-6 Astra performance on benchmarks like Terminal-Bench 4.0 and AutomationBench-AA while offering improved cache hit discounts.

  19. 019Hacker NewsSEP · 22English

    Context compaction, measured: FutureOS vs. Codex vs. OpenCode

    A comparative study tested three context compaction strategies—FutureOS, OpenCode, and Codex—on their ability to retain information from agent sessions. FutureOS retained 83% of queryable information, significantly outperforming OpenCode (47%) and Codex (38%), with the key difference being that FutureOS preserves assistant prose while others compress it away. The analysis reveals that tool output dominates context volume but is rarely referenced, while the sparse assistant text is the primary source of follow-up questions.

    By FutureOS
  20. 020Hacker NewsSEP · 21English

    Transformers Explained Visually

    Transformers are a neural network architecture introduced in 2017 that power modern AI models like GPT, Llama, and Gemini. They use self-attention mechanisms to predict the next token in sequences and process text through embedding, transformer blocks with attention layers, and output probability layers.

    By aray07