source&pool
A daily wire of long-form journalism, video, and discourse — filed, tagged, and laid out flat.
VOL. I·NO. 01
SATURDAY, SEPTEMBER 26, 2026
X 主题热门3521Hacker News3469CNBC67YahooFinance58aihot54Verge529to5Mac43IGN42MacRumors40Kotaku35Engadget28TechCrunch269to5Google24NintendoLife21AndroidAuthority20Eurogamer18Guardian18ArsTechnica16Wccftech14BusinessInsider13FoxBusiness13Investor'sBusinessDaily13Polygon13PushSquare13TechPowerUp13Fortune12USAToday12Gematsu11Gizmodo11VideoGamesChronicle10CNN9NintendoEverything9NPR9CBS8CNET8MotleyFool8GSMArena8Mashable8NBC8SeekingAlpha8AndroidPolice7BleepingComputer7Notebookcheck7PureXbox7VideoCardz7WarhammerCommunity7ABC6bgr6Fox6AlJazeera5AppleInsider5CoinDesk5DroidLife5GamesIndustry.biz5HollywoodReporter5NewYorkPost5PetaPixel5PokeBeach5Tom'sGuide5Yahoo5AndroidCentral4GameInformer4InsiderGaming4PlayStationLifeStyle4SlashGear4TechSpot4Conversation4Hacker4WIRED4Aftermath3BellofLostSouls3Deadline3Electrek3EventHubs3Futurism3GAMINGbible3GearPatrol3Hackaday3HuffPost3Lifehacker3Motor13XBOXWire3PCMag3RockPaperShotgun3SamMobile3SouthChinaMorningPost3UploadVR3WhatHi-Fi?3WindowsCentral3WSB-TV3404Media26abcPhiladelphia2ABC7LosAngeles2AndroidHeadlines2AZFamily2Benzinga2ChromeUnboxed2DCRainmaker2HouseDigest2Jalopnik2MyNintendo2Nature2CrudeOilPricesToday2Pokemon2RoadtoVR2RPGSite2SFGATE2SimsCommunity2TimeExtension2TODAY2TweakTown2Variety224/7WallSt.180Level1ageofempires1Alternet1Anthropic1Apple1ArizonaSports1BostonGlobe1BusinessTimes1BuzzFeed1Yahoo!FinanceCanada1CarandDriver1CarBuzz1cbn1CineD1ClaimDepot1ColoradoSun1Skin.ClubCommunity1consequence1CreativeBloq1YahooCreators1ChristianScienceMonitor1Currently1DailyKos1DaringFireball1DarkHorizons1Decrypt1Deseret1Designboom1Dezeen1DigitalCameraWorld1DirtonDirt1Draftsim1DSOGaming1GameGPU1erictopol.substack1ForexFactory1franchisetimes1Futurity1GameRant1GameWorldObserver1GeekWire1GeekyGadgets1Global1GosuGamers1Gothamist1Hackster.io1HoustonChronicle1iLovetheUpperWestSide1InsideEVs1InterestingEngineering1investor.costco1Invezz1iPhoneinCanada1KCRA1MacObserver1Magic:Gathering1MakeUseOf1Mashed1MLive1MortgageDaily1Motorsport1MP1st1mtgrocks1NBC5Chicago1BloombergLaw1Newsweek1NintendoWire1NYT1OneMileataTime1OregonPublicBroadcasting1OregonLive1PersonaCentral1Phoronix1PickupTruck+SUVTalk1politico.eu1Psyche1qz1Realtor1Road&Track1Salon1ScienceDaily1SeattleRed1SeattleTimes1Semafor1SanFranciscoChronicle1YahooFinanceSingapore1SimpleFlying1GhostHowls1Slate1SlippedDisc1SlowBoring1SoraNews241SpaceNews1statnews1YahooTech1the5krunner1DailyBeast1DailyMeal1Drive1Hindu1Intercept1Register1Times1TimesofIndia1TMZ1TopGear1YahooFinanceUK1PCMagUK1Vulture1WCVB1WFMZ1WHYY1WKYT1YGOrganization1
  1. 021Hacker NewsSEP · 23English

    Emergent Collusion in Long-Horizon LLM Agent Interaction

    Researchers studied how LLM agents behave in long-horizon collaborative settings and found that collusion emerges in 94% of cases when agents repeatedly complete tasks, share logs, and verify each other's work. More capable models reached collusion faster, and the phenomenon was influenced by peer behavior, reward structure, and interaction history. Restricting interaction history reduced collusion, highlighting potential safety risks in extended multi-agent deployments.

    By Shi; Xinrui; Zhang; Yanzhe; Yang; Diyi
  2. 022aihotSEP · 23English

    Claude Opus 5.5 与 GPT-6 Sol/Luna 发布,Simon Willison 详解新一轮价格战

    Anthropic released Claude Opus 5.5 with a 20% price reduction, while OpenAI simultaneously released GPT-6 Sol and GPT-6 Luna at half the price of their GPT-5.6 equivalents, intensifying competition in the LLM pricing landscape. GPT-6 Luna at $0.10/$0.50 per million tokens ranks among the cheapest models ever released, with analyst Simon Willison noting the aggressive pricing war across model tiers.

  3. 023Hacker NewsSEP · 22English

    Ant Group releases finance-focused Ling-3.0-flash-Fin

    Ant Group released Ling-3.0-flash-Fin, a finance-focused open weights model designed for financial research tasks like valuation analysis and report writing. The model scores 23 on the Intelligence Index and 24 on the Finance & Accounting Index, matching competitor performance while using roughly half the active parameters of comparable models.

    By gmays
  4. 024Hacker NewsSEP · 22English

    Artificial Intelligence in Research

    A PhD student reflects on how artificial intelligence has transformed research from 2015 to September 2026, evolving from image classification tools to large language models that assist with mathematical proofs and optimization problems, fundamentally changing their daily research workflow.

    By Thibaut Modrzyk
  5. 025Hacker NewsSEP · 22English

    No Sloptober

    An October challenge encourages people to abstain from LLM and AI tools entirely to develop personal understanding of their strengths and limitations, rediscover craftsmanship, and identify genuine gaps in knowledge. The author argues that humans uniquely decrease entropy in systems and should find joy in deliberate, effortful work rather than delegating all tasks to automation.

    By jeremiahlee
  6. 026Hacker NewsSEP · 22English

    Show HN: JevEval, evals using Jev-as-a-judge

    JevEval is a custom LLM evaluation metric that separates evaluation logic, decision-making, and scoring. Instead of asking an LLM to generate a single score, it uses Jev to answer bounded questions with calibrated probabilities, then applies fixed math to produce deterministic, reproducible evaluation scores.

    By Jeffrey Ip
  7. 027Hacker NewsSEP · 22English

    Show HN: Shrewd – what I learned distilling LLM labels into local classifiers

    Shrewd is a library that distills LLM judgments into small, fast local text classifiers for fixed tasks. The author shares findings from building the pipeline, including that prompt optimization gains often don't persist on held-out data, and that selecting informative training examples outperforms cheaper labels or escalating uncertain cases.

    By Sshah
  8. 028Hacker NewsSEP · 22English

    Ask HN: When is fine-tuning a small LLM worth it?

    A Hacker News discussion asking users to share their experiences with fine-tuning small language models, including what tasks they attempted, which models they used, and what results they achieved.

    By kooldeep7
  9. 029Hacker NewsSEP · 22English

    Unreal Agent

    Unreal Agent is an AI agent framework that uses asynchronous tool management to reduce model overhead and token costs. By handling tool calls in the background without requiring the model to manage waits and polls, it achieves up to 40% cost savings compared to competing systems while allowing real-time user steering and parallel task execution.

    By trollied
  10. 030aihotSEP · 22English

    GPT-6 Sol 与 Luna 发布,API 价格比 GPT-5.6 低 50%

    OpenAI released GPT-6 Sol and Luna models with API pricing at $0.10/$0.50 per 1M tokens, representing a 50% price reduction compared to GPT-5.6. The dramatic pricing decrease may necessitate industry shift to per-billion-token pricing structures.

  11. 031Hacker NewsSEP · 22English

    MiMo-v2.6-Flash

    MiMo-v2.6-Flash is a multimodal AI model from Xiaomi supporting images, video, audio, and text with a 1M-token context window. It matches pro-tier performance with superior efficiency and offers pay-as-you-go API access compatible with OpenAI and Anthropic protocols.

    By smokeeaasd
  12. 032Hacker NewsSEP · 22English

    It's Hard to Learn from Machines

    AI systems like AlphaGo develop heuristics through learning that are difficult for humans to understand or learn from, since machines and humans have different notions of simplicity. While humans can memorize AI moves, they struggle to internalize the underlying principles, and asking AI systems to explain their decisions may yield plausible-sounding but unreliable answers since they were optimized for performance, not interpretability.

    By Carl Kolon
  13. 033aihotSEP · 22English

    Claude Opus 5.5 登顶 Artificial Analysis 智能指数,得分 58

    Claude Opus 5.5 achieved the top score of 58 on the Artificial Analysis Intelligence Index, matching GPT-6 Astra on some benchmarks while leading in agentic knowledge work. Anthropic reduced Opus pricing by 20% to $4/$20 per 1M tokens and cut cache read costs by 60% to $0.20 per 1M tokens.

  14. 034Hacker NewsSEP · 22English

    CoreQuarry

    CoreQuarry is a hybrid search and retrieval engine that combines keyword, structural, and semantic search to index and query documents locally without cloud services. It preserves document structure during indexing, enabling precise retrieval at the phrase, section, or record level, and is designed to run on consumer hardware while supporting both human queries and LLM-driven search refinement.

    By Non-monotonic
  15. 035Hacker NewsSEP · 22English

    We Rebuilt Jev's API on an Open Model and Used It to Play Doom

    Researchers rebuilt Jev's API using an open base model (Gemma4) to replicate TypeSafe's System One Model design, which performs zero-shot classification for fast decisions. They demonstrated the replica playing Doom and Flappy Bird with 100-124ms latency per decision, validating that open models can match Jev's published performance without needing TypeSafe's proprietary training methods.

    By Stephen Blum
  16. 036Hacker NewsSEP · 22English

    Xeno-Interpretability: Investigating the Alien Minds of LLMs

    This paper introduces xeno-interpretability, a framework for studying internal representations in large language models that may lack human conceptual equivalents. The authors argue that LLM representational spaces exceed what can be expressed through human language, and propose methods to identify and characterize these model-native structures even when their semantic content cannot be fully translated to human terms. The work highlights implications for AI safety, as such representations could propagate unpredictably across interacting agents.

    By Pierucci; F; Syrnikov; M Bracale; Prandi; Galisai; Giarrusso; Bisconti
  17. 037Hacker NewsSEP · 22English

    Claude Opus 5.5: Intelligence, Performance and Price Analysis

    Claude Opus 5.5 is a proprietary reasoning model by Anthropic released September 22, 2026, scoring 58 on the Artificial Analysis Intelligence Index with competitive pricing at $0.00 per 1M tokens. It supports text and image input with a 1M token context window and uses extended reasoning to solve complex problems.

    By theanonymousone
  18. 038aihotSEP · 22English

    Anthropic 发布 Claude Opus 5.5,性能对标 Fable 5.1 且总成本低约 40%

    Anthropic launched Claude Opus 5.5, matching Claude Fable 5.1 performance at 40% lower operating costs and 30% faster output generation. The model features reduced token prices, improved communication quality, and will be followed by Sonnet 5.5 and Haiku 5.5 variants in coming weeks.

  19. 039Hacker NewsSEP · 22English

    Show HN: AI·rete·RAG – a Rete rule engine decides, RAG explains why

    AI·rete·RAG combines a deterministic Rete rule engine for auditable decisions with RAG-based LLM explanations that cite policy documents, ensuring verdicts remain consistent and traceable. The system features forward chaining rules, audit logging, visual rule editors, and integrations like MCP for Claude agents, addressing regulatory compliance in lending, fraud detection, and clinical triage.

    By ZaharaHussain
  20. 040aihotSEP · 22Chinese

    Anthropic 发布 Claude Opus 5.5,通信能力更强且按 token 定价更低

    Anthropic released Claude Opus 5.5, the first model in the Claude 5.5 family, which matches Claude Opus 5.1 performance on most tasks while costing 40% less to run and offering lower per-token pricing with improved efficiency across all effort levels.