source&pool
A daily wire of long-form journalism, video, and discourse — filed, tagged, and laid out flat.
VOL. I·NO. 01
TUESDAY, SEPTEMBER 29, 2026
X 主题热门3639Hacker News3616CNBC73aihot72YahooFinance67Verge59IGN489to5Mac38MacRumors36Engadget349to5Google33Kotaku30TechCrunch29AndroidAuthority28PushSquare23Eurogamer20NintendoLife19ArsTechnica18Guardian18Investor'sBusinessDaily17Mashable15TechPowerUp15FoxBusiness14Wccftech14Polygon13SeekingAlpha13Gematsu12BusinessInsider11Fortune11Gizmodo11NBC11bgr10CNN10GSMArena9NPR9VideoGamesChronicle9AlJazeera8AndroidPolice8GameGPU8PureXbox8USAToday8CNET7DroidLife7MotleyFool7Fox7XBOXWire7NewYorkPost7WIRED7BleepingComputer6CBS6GameInformer6GamesIndustry.biz6Jalopnik6NintendoEverything6Hacker6AndroidCentral5AppleInsider5HollywoodReporter5Tom'sGuide5VideoCardz5Yahoo5ABC4Draftsim4GAMINGbible4InsiderGaming4NYT4CrudeOilPricesToday4PokeBeach4UploadVR4WarhammerCommunity4404Media3Aftermath3Electrek3EventHubs3Lifehacker3Nature3PCMag3PetaPixel3PlayStationLifeStyle3SouthChinaMorningPost3Space3YahooTech3TechSpot3Conversation3WhatHi-Fi?324/7WallSt.2Autonocion2AZFamily2BostonGlobe2BuzzFeed2ChromeUnboxed2Currently2DCRainmaker2Deadline2Futurism2GameRant2GearPatrol2Hackaday2iLovetheUpperWestSide2MP1st2mtgrocks2Notebookcheck2Pokemon2politico.eu2RoadtoVR2RockPaperShotgun2RPGSite2SamMobile2SFGATE2Yahoo2SimpleFlying2SimsCommunity2SlashGear2Slate2Register2TimeExtension2TODAY2WindowsCentral2ynetnews26abcPhiladelphia180Level1WXLV1ageofempires1airlive1AJC1AndroidHeadlines1Anthropic1Apple1AVClub1AviationWeek1Benzinga1BoingBoing1Yahoo!FinanceCanada1YahooLifestyleCanada1CarandDriver1CineD1CnEVPost1CoinDesk1comicbookmovie1Skin.ClubCommunity1consequence1CreativeBloq1YahooCreators1DailyKos1DaringFireball1Deseret1Designboom1Dezeen1DigitalCameraWorld1DigitalFoundry1DSOGaming1DiarioAS1Euronews1Finbold1ForexFactory1FOX13Seattle1franchisetimes1Futurity1GameFile1GameWorldObserver1garymarcus.substack1GeekWire1GeekyGadgets1GosuGamers1Gothamist1Hackster.io1HuffPost1Independent1InsideEVs1InterestingEngineering1investor.costco1Invezz1I/OFund1iPhoneinCanada1KCRA1KOMO1LosAngelesTimes1MacObserver1Magic:Gathering1MakeUseOf1Mashed1MLive1MortgageDaily1Motor11Motorsport1MyNintendo1MyNorthwest1NBCBayArea1NBC5Chicago1NBC7SanDiego1BloombergLaw1Newser1SemiAnalysis1Newsweek1NintendoWire1Nokiamob1OregonPublicBroadcasting1OregonLive1PersonaCentral1Phoronix1PickupTruck+SUVTalk1Psyche1QuantaMagazine1Realtor1RichmondTimes-Dispatch1Richmonder1Road&Track1ScienceDaily1ScreenRant1SeattleRed1SeattleTimes1SanFranciscoChronicle1YahooFinanceSingapore1GhostHowls1SlowBoring1SoraNews241SpaceNews1statnews1svg1TampaBayTimes1the5krunner1Autopian1DailyBeast1Intercept1NextWeb1https://tipswatch.com/1TMZ1TopGear1TweakTown1YahooFinanceUK1PCMagUK1Variety1Vulture1WCVB1WFMZ1WHYY1WrestlingInc.195.5WSB1YGOrganization1
  1. 001Hacker NewsSEP · 29English

    Uncensored and Offensive Security AI Models Benchmark

    A curated benchmark of open-weight large language models fine-tuned for offensive security, penetration testing, and red team operations. The list includes 27 models with varying sizes and capabilities, sourced from HuggingFace, alongside technical descriptions of abliteration and fine-tuning methods used to remove alignment constraints.

    By JoasASantos
  2. 002Hacker NewsSEP · 28English

    MoE Analysis Qwen[3.5|3.6]35B-A3B

    Analysis of Mixture-of-Experts routing in Qwen 3.5 and 3.6 35B models using MT-Bench prompts, examining which experts are activated per token and the router's confidence levels across MoE layers.

    By Gurpreet Singh Laller; Opus
  3. 003Hacker NewsSEP · 28English

    Holo4: Powering generalist computer-use agents

    Holo4 is a new series of agentic models (27B dense and 35B-A3B MoE) designed for computer-use tasks that interact with software through GUIs, code, APIs, and MCP. Trained via supervised and reinforcement learning on diverse environments, it scores competitively with frontier models on academic benchmarks like OSWorld 2.0 while operating at significantly lower cost and parameter count.

    By Tony Wu; Maxime Theillard; Frederic Renard; Vincent Coyette; Emrick Sinitambirivoutin; Avshalom Manevich; Antonio Loison; Antoine Bonnet; Maxime Langevin; Aleix Cambray; Léonard Benedetti; Mats L Richter; Michael Eickenberg; Sławek Mucha; Matthias Brunel; Daniel Beechey
  4. 004Hacker NewsSEP · 28English

    Orcarouter/OrcaSAQ-2-27B

    OrcaSAQ2 27B is a 3-bit quantized version of Qwen3.8-27B that compresses the model from 54 GB to 12.3 GB while maintaining 93.2% token-level agreement and only +0.02% perplexity increase. Optimized for long-horizon agent tasks like coding, tool use, and reasoning, it enables deployment on 16 GB GPUs with strong performance on benchmarks like SWE-bench and Terminal-Bench.

    By handfuloflight
  5. 005Hacker NewsSEP · 28English

    Jeff – Jev-compatible 0.8B decision models, trained at home, ~30 ms

    Jeff is a small 0.8B decision model fine-tuned from Qwen and Gemma for fast zero-shot classification tasks, making calibrated probability judgments between user-defined options in ~22-28ms. Trained entirely on local hardware using synthetic data, it matches or exceeds larger models on classification benchmarks while acknowledging weaker reasoning performance on complex tasks.

    By Firelex
  6. 006Hacker NewsSEP · 28English

    Small Decisions: Engineering a Leading Model

    An AI engineer documents building Hobson, a 2-billion-parameter calibrated classifier model by fine-tuning Qwen3.5-2B with a custom pointer head and LoRA adapter. The model achieves top performance in its size range on the JevBench leaderboard, demonstrating superior calibration and accuracy compared to existing alternatives through careful architectural choices and self-distillation training techniques.

    By Marc Brooker
  7. 007Hacker NewsSEP · 28English

    It Was the Harness, Not the Model

    A developer tested five coding agents on the same local model and frozen test suite, finding that 90% of failures stemmed from harness problems rather than model limitations. Contrary to expectations, using a larger or less-quantized model did not fix these harness-related issues, demonstrating that tooling quality matters more than raw model capability.

    By aray07
  8. 008Hacker NewsSEP · 28English

    I got 2.2x more tokens per second from llama.cpp on Intel Arc

    A developer benchmarked llama.cpp on an Intel Arc-equipped laptop and found that disabling CPU offload for mixture-of-experts layers achieved 2.2–2.3x faster performance, though it requires careful memory management. Key optimizations included using speculative decoding with n=2 and increasing batch size to 2048 for long prompts, while thread count and CPU governor had negligible impact.

    By Luigi
  9. 009Hacker NewsSEP · 28English

    LLM capabilities can transfer through unrelated text

    A research paper demonstrates that language model capabilities can transfer to unrelated tasks through post-training artifacts. The release includes reproducible code and frozen training data for experiments using the Qwen2.5-1.5B model on HumanEval+ benchmarks, with detailed instructions for verification and replication across CPU and GPU environments.

    By Myboker
  10. 010Hacker NewsSEP · 27English

    How Jev works: calibrated decision models

    Jev is a decision model from TypeSafe AI that selects among predefined options by scoring their likelihood as text completions, then normalizing scores into probabilities—without generating text. Experiments on Qwen2.5 models show scoring is 7–54× faster than text generation, with comparable accuracy, though performance degrades as option counts increase. Small models can be fine-tuned efficiently to improve calibration and accuracy on specific tasks.

    By Victor Dibia
  11. 011Hacker NewsSEP · 27English

    Life Forge – Autonomous Flight Simulator for AI Agents

    Life Forge is an autonomous flight simulator for testing AI agents using co-evolutionary adversarial red-teaming and 3D MAP-Elites algorithms. It dynamically stress-tests frontier models like Claude, GPT-4o, and Qwen in simulated enterprise environments with realistic perturbations such as price volatility, supply scarcity, and prompt injections to expose failure modes before production deployment.

    By Zariffromlatif
  12. 012Hacker NewsSEP · 27English

    Glyd – Run LLMs in 33% less GPU memory, bit for bit

    Glyd is a lossless AI compression technique that stores open-source LLM weights in 11 bits instead of 16, reducing GPU memory usage by 33% while maintaining bit-for-bit accuracy. The method decodes weights directly in GPU matrix operations without rounding, enabling the same models to run on less hardware, often faster, with negligible impact on model outputs.

    By surya-koritala
  13. 013Hacker NewsSEP · 27English

    Jev and the Return of AI/ML Engineering

    Jev markets itself as a specialized 'System One' model with calibrated probabilities, structured outputs, and low latency/cost, but most capabilities can be replicated with existing open-source models like Qwen or DeepSeek. While Jev shows potential for rubrics and preference modeling, its core calibration claims appear overstated based on empirical tests showing high calibration errors across datasets.

    By Han Lee
  14. 014Hacker NewsSEP · 27English

    Jev model free online with direct API

    Sifty is a free online API for text and image classification using custom labels, powered by the open-source Qwen3.5-4B model. It offers 1,000 free units per IP daily, supports HTTP API calls and Python SDK integration, and can be self-hosted. The service does not store images or require authentication.

    By ralphlaur
  15. 015Hacker NewsSEP · 27English

    I expect AI replication incidents by 2027

    An analyst predicts a major autonomous AI replication incident in the wild by end of 2027, driven by open models running efficiently on consumer hardware and the capability gap between frontier and open models narrowing. Such AI worms could enable deniable false-flag operations between rival AI labs and states, with the US and China each holding distinct advantages in orchestrating or defending against such attacks.

    By Reworr R
  16. 016Hacker NewsSEP · 27English

    What reversing, modernising old games tells us about the economic impact of AI

    A software engineer used AI models like Qwen and GPT-6 Astra to reverse engineer and modernize the 1989 game War of the Lance for web browsers, then upgraded it with new gameplay features. The work demonstrates AI's accelerating capability in code synthesis, 3D asset generation, and reverse engineering, enabling a single person to accomplish complex game porting tasks in hours rather than weeks.

    By Georg Zoeller
  17. 017Hacker NewsSEP · 26English

    Modern LLMs have tiny GPTs hidden inside them

    A researcher tested whether modern LLMs contain internal models of other LLMs by having Qwen complete text started by GPT-2, comparing whether Qwen's continuations resembled GPT-2's own continuations more than Qwen's natural output. The experiment used headlines from September 2026 (outside training cutoffs) and varied generation lengths to measure textual overlap between conditions.

    By Paras Chopra
  18. 018Hacker NewsSEP · 26English

    Typed-lm: a Rust jev open source alternative

    typed-lm is an open-source Rust framework that converts large language models like Llama and Qwen into typed semantic-routing APIs, enabling deterministic inference with millisecond latency by returning structured outputs (booleans, choices, scores) instead of generated text. It includes a trainer for adapter-based specialization and supports quantization for efficient deployment on GPU and CPU.

    By Neurono-Ml
  19. 019Hacker NewsSEP · 26English

    Getting Deeper into Local Inference

    A developer describes transitioning to local LLM inference using Qwen3.8-27B on an RTX 5090 laptop, finding it sufficient for personal Q&A, coding, and sysadmin tasks. Hardware upgrades and model improvements now make local deployment viable, though economically inferior to cloud providers, with speed and privacy being primary motivations.

    By thombles
  20. 020Hacker NewsSEP · 26English

    Custom Models in Oh My Pi: vLLM, Llama.cpp, SGLang and More

    Oh My Pi (omp) supports custom models through local inference servers like vLLM, llama.cpp, SGLang, and gateways by configuring ~/.omp/agent/models.yml. Recent updates require renaming custom providers and adding Qwen template settings for reasoning effort compatibility.

    By Doug Calobrisi