MikeTrendsTrends right now

search

large language models

Trends

  1. 1
    Anthropic Launches New AI Evaluation and Optimization Tools●Anthropic Launches Tools for Reliable AI Evaluations and Optimization𝕏xSETechnologyAI6011 d ago

    Anthropic has released a set of tools designed to help developers run more reliable AI evaluations and optimize their models. The tools aim to make it easier to measure model performance, compare versions, and improve output quality in production systems. The announcement is drawing attention from developers and AI industry watchers tracking how companies test and refine large language models.

  2. 2
    OpenAI and Microsoft researchers warn of AI 'doom loop' consuming the web▼‘Doom Loop’: OpenAI and Microsoft Admits LLMs Are Destroying the Web and Built on TheftMmastodonTechnologyAI851 d ago

    Researchers at OpenAI and Microsoft have published a paper describing a 'doom loop' in which large language models, trained on data scraped largely without consent from human creators, degrade the open web and eventually poison their own training data. The paper reportedly acknowledges that people will come to see the models' wholesale ingestion of creative work as an unprecedented act of theft, reigniting debate over copyright and the sustainability of generative AI.

  3. 3
    China's Cheap Open AI Models Now Dominate Global Developer Use▼China's Cheap, Open AI Models Now Power Most Global Developer Traffic, Alarming Washington✉newsTechnologySoftware2 d ago

    Chinese-developed artificial intelligence models, offered at low cost and with open access, now account for the majority of traffic from developers worldwide building on large language models. The trend is drawing concern in Washington, where officials worry that widespread reliance on Chinese AI infrastructure could give Beijing strategic and technological leverage over the global software ecosystem.

  4. 4

    A new study is probing whether artificial intelligence systems could experience pain, one of the hardest questions in debates over machine consciousness. The research feeds into a wider argument among scientists and ethicists about whether large language models have any form of subjective experience, and what that would mean for how such systems should be treated and regulated.

  5. 5

    Ollaya is a project being discussed on Hacker News, described as 'Ollama for open-source, Jev-style decision models'. The framing suggests a tool that makes decision-making models as easy to run locally as Ollama made large language models, though the single post title gives little detail. With 537 likes and a high rank, commenters appear interested in the analogy to Ollama, but the posts collected do not explain what the tool actually does or why it is generating attention.

  6. 6
    Florida sues OpenAI, calling advanced AI a public nuisance●TL;DR: Florida argues that OpenAI's development of large language models poses a significant risk to civilization, filinMmastodonWorldLaw & Courts71 d ago

    Florida has filed a legal challenge against OpenAI, arguing that the company's development of large language models poses a significant risk to civilization. The state is seeking to halt further progress by framing the work as a public nuisance, an unusual legal theory for AI regulation. The move marks one of the most aggressive state-level attempts to restrain AI development and is drawing attention from legal and tech observers.

  7. 7
    Chinese 'robot brain' startup eyes ChatGPT-style breakthrough next year▼Chinese 'robot brain' startup sees ChatGPT-style breakthrough as soon as next year✉newsTechnologyRobotics2 d ago

    A Chinese robotics startup says it could achieve a ChatGPT-style breakthrough in robot intelligence, often called a 'robot brain', as soon as next year. The claim suggests rapid progress toward general-purpose AI for robots, allowing machines to learn and adapt the way large language models do, and would mark a major milestone in China's push to compete in embodied AI.

  8. 8
    Stolen AI credentials fuel underground LLM proxy market▼Stolen AI credentials feed growing LLM proxy economy✉newsBusinessEconomy1 d ago

    Security researchers report that stolen AI credentials are being sold and traded to power a growing black market of LLM proxies, where criminals resell access to paid large language model services at cut-rate prices. The trade lets buyers run AI workloads on accounts billed to victims, exposing companies to unexpected costs and data risks as adoption of AI tools accelerates.

  9. 9
    Supersonic Labs Releases Julia 1, a CPU-Friendly Open Decision Model●Supersonic Labs Releases Julia 1: A 144.3M-Parameter Open Decision Model That Runs on a CPU✉newsTechnologySoftware2 d ago

    Supersonic Labs has released Julia 1, an open decision model with 144.3 million parameters that is small enough to run on a standard CPU. Unlike large language models, decision models are built for making choices and taking actions rather than generating text. The low hardware requirement makes the model accessible to developers without expensive GPU infrastructure, which is drawing attention in the AI community.

  10. 10
    Microcontrollers now run a diffusion model and 289M-parameter LLM▼Microcontrollers now run a diffusion model and 289M LLM✉newsTechnologySoftware21 min ago

    Tiny microcontroller chips, traditionally limited to simple embedded tasks, can now run a diffusion model for image generation and a compact 289-million-parameter large language model. The news, highlighted by Adafruit and Open Source For You, points to rapid progress in on-device AI, letting small, low-power hardware perform generative tasks without cloud servers. Enthusiasts are discussing what this means for smart devices, robotics and offline AI applications.

  11. 11
    Vincent Says General-Purpose Robots Could Arrive Within Two Years▼Vincent: General-Purpose Robots Arrive Within Two Years — and LLMs Are Driving Them✉newsTechnologyRobotics3 d ago

    Vincent claims that general-purpose robots may reach practical deployment within two years, driven largely by large language models acting as their reasoning and control layer. The prediction frames LLMs as the key breakthrough allowing robots to understand instructions and adapt to unfamiliar tasks. The claim is drawing attention as robotics firms race to combine AI models with hardware.

  12. 12

    TensorFlow, Google's open-source machine learning framework, is trending on GitHub this week. The repository provides tools for building and training machine learning models, and is written largely in C++ with interfaces for Python and other languages. The posts visible are simply links to the repository with its standard description, so there is no specific release, announcement, or discussion evident from the snippets. Trending likely reflects renewed attention from developers, but the exact trigger is not clear from the posts.

  13. 13
    Roboharm study tests whether robots refuse unsafe instructions▼Roboharm: Do frontier robot policies refuse unsafe instructions?YhnTechnologyRobotics601 d ago

    A project called Roboharm is asking whether frontier robot policies actually refuse unsafe instructions. The work examines how AI systems controlling robots respond to harmful commands, a safety question increasingly urgent as AI models are deployed in physical robotics. Discussion is focused on how well current safeguards carry over from language models to embodied systems.

  14. 14
    Landemore and Tang call for global citizen commission to oversee AI●Vor kurzem ist ein Artikel von Hélène Landemore und Audrey Tang im Noema Magazin erschienen. Es wird die Forderung gesteMmastodonTechnologyAI12 d ago

    Political theorist Hélène Landemore and Taiwan's former digital minister Audrey Tang have published an article in Noema Magazine calling for the development of large language models to be placed under the supervision of a rotating global commission of 1,000 members, selected like a citizens' assembly. The proposal links democratic sortition with AI governance and is drawing attention in discussions about who should control powerful AI systems.

  15. 15
    Amit Sahai argues the world needs far more mathematicians▼Intriguing 🤔: “We’re Gonna Need A Lot More Mathematicians”, Amit Sahai via Terence Tao ( https:// terrytao.wordpress.comMmastodonTechnologyAI22 d ago

    Cryptographer Amit Sahai has published a guest essay titled 'We're Gonna Need A Lot More Mathematicians' on Fields Medalist Terence Tao's blog. The piece, also circulating on Hacker News, argues that mathematics expertise will be in higher demand as AI tools advance, prompting discussion among researchers about how large language models change the value and supply of mathematical talent.

  16. 16
    Clip of GPT-6 Astra steering Unitree G1 humanoid goes viral●GPT-6 Astra pilots a Unitree G1 humanoid in an unseen room — Reddit's verdict on the viral demo A 52-second silent clipMmastodonTechnologyRobotics12 d ago

    A 52-second silent clip showing a Unitree G1 humanoid robot tidying a room it had never seen before, attributed to GPT-6 Astra, has drawn more than 1,400 upvotes on r/singularity. Commenters are debating whether the demo genuinely shows a large language model controlling the robot in an unfamiliar space, or whether the footage is overhyped or staged. The exchange reflects growing public scrutiny of claims linking frontier AI models to physical robotics.

  17. 17
    AI model Jev beats Pokémon Red in under a week●Developer says AI decision model Jev beat Pokémon Red in under a week — non-LLM engine succeeds where traditional chatbots stalled for months, but Claude Opus 5 coached the model through its dead ends✉newsTechnologyAI2 d ago

    A developer says Jev, a non-LLM AI decision model, has completed Pokémon Red in under a week, a feat that reportedly stalled traditional chatbot-based attempts for months. According to the report, Claude Opus 5 acted as a coach, helping Jev work through dead ends during the run. The result is being discussed as evidence that specialized decision engines can outperform large language models on structured, long-horizon tasks like game completion.

  18. 18
    Why general-purpose LLMs may win at robotics▼Why general-purpose LLMs may win at robotics — Waddle Labs and RoboCurve✉newsTechnologyRobotics1 d ago

    Waddle Labs and RoboCurve are the focus of a new analysis arguing that general-purpose large language models could outperform specialised systems in robotics. The piece contends that broad foundation models trained on diverse data may adapt better to physical tasks than narrowly built robots, a view that challenges a common assumption in the field and is drawing attention among robotics and AI investors.

  19. 19
    UN University proposes governance framework for LLM agent simulations●From Plausible Agents to Accountable Simulation: A Technical and Governance Framework for LLM-Enabled Agent-Based Modelling✉newsTechnologyAI18 h ago

    United Nations University researchers have published a technical and governance framework for using large language models in agent-based modelling, titled 'From Plausible Agents to Accountable Simulation'. The work addresses how LLM-enabled simulations, which can produce realistic-seeming artificial agents, can be made verifiable, transparent and accountable when used for research and policy analysis. It proposes standards for evaluating whether simulated agent behaviour is plausible and for governing the use of such models.

  20. 20
    Alpine Linux contributors vote against banning LLM-generated code●@ gildilinie # Alpine # Linux had a vote among core contributors, and similar to debian, the majority wasn't in favor ofMmastodonTechnologyAI02 d ago

    Alpine Linux held a vote among its core contributors on whether to ban code written with large language models, and the majority voted against a ban. The result mirrors an earlier vote in the Debian project, which also declined to prohibit LLM-generated code. The decision was recorded in the Alpine council's meeting minutes and is being discussed by open source developers weighing how much AI assistance to allow in volunteer-built distributions.

  21. 21
    Non-LLM AI model beats Pokémon Red in under a week●Developer says Jev decision model beat Pokémon Red in under a week — non-LLM engine succeeds where traditional chatbots stalled for months, but Claude Opus 5 coached the model through its dead ends✉newsTechnologyAI2 d ago

    A developer says a decision-model system called Jev beat Pokémon Red in under a week, succeeding where LLM-based agents have stalled for months. The engine itself is not a language model, but Claude Opus 5 reportedly acted as a coach, helping it past dead ends. The claim has drawn attention from AI watchers who see it as a counterpoint to the belief that large language models are the best path to autonomous game-playing agents.

  22. 22
    Solus Linux Adopts Formal AI Contribution Policy●Linuxiac: Solus Linux Adopts Formal AI and LLM Contribution Policy https:// linuxiac.com/solus-linux-adopt s-formal-ai-aMmastodonTechnology13 d ago

    Solus Linux, the independent Linux distribution, has introduced a formal policy governing contributions created with AI and large language models. The move, reported by Linuxiac, sets clear rules for how AI-assisted code is handled in the project. It reflects a broader debate in open-source communities about how to manage the growing volume of AI-generated submissions to volunteer-maintained software projects.

  23. 23
    Pope Leo urges humanity to 'remain human' in age of AI▼Cyborgs, centaurs and LLeMmings: Why Pope Leo calls us to 'remain human'✉newsTechnologyRobotics1 d ago

    Pope Leo is calling on people to 'remain human' amid rapid advances in artificial intelligence and robotics. A National Catholic Reporter commentary explores the theme through imagery of cyborgs, centaurs and 'LLeMmings' — a play on large language models — reflecting on where human dignity fits as machines take on more roles once reserved for people, and urging discernment rather than blind adoption of new technologies.

  24. 24
    Viewers compare Pluribus's alien conversations to AI chatbots●Rewatched Pluribus and was struck by 1. How much the conversations with the Others reminds me of LLMs: the world’s knowlMmastodonCultureTelevision38 h ago

    Pluribus, the Apple TV sci-fi series, is drawing fresh attention from viewers who see parallels between the show's collective alien minds, known as the Others, and today's large language models: a repository of the world's knowledge delivered in a charming yet unsettling way. Alongside the AI comparisons, fans are voicing impatience for the confirmed second season to arrive sooner.

  25. 25

    Pirate Face is a site or service being discussed on Hacker News under a headline claiming it 'rescues LLM models from deletion'. The post suggests it offers a way to preserve or recover large language models that might otherwise be removed, but the linked evidence gives no detail about how it works or who is behind it. The platform labels it as medicine, which does not obviously match the topic, and commenters' reactions are not visible, so sentiment and specifics are unclear from the posts alone.

  26. 26
    How to Query and Compare Multiple LLMs on Linux●How to Query and Compare Multiple LLMs on Linux # ai # anthropic # api # chatgpt # claude # large_language_models_ (llmsMmastodonTechnologyAI114 h ago

    A guide circulating among Linux users explains how to query and compare several large language models, including Claude, ChatGPT, DeepSeek, Grok and GLM, from the command line using Python and API aggregators. It walks through setting up access to multiple providers so outputs can be reviewed side by side, aimed at developers choosing between AI services or testing models on their own machines.

  27. 27
    Solus Linux adopts official policy on AI-generated contributions●Solus Linux adoptă o politică oficială privind contribuțiile generate de AI și LLM https:// linuxforeducation.blogspot.cMmastodonTechnologyAI12 d ago

    The Solus Linux distribution has adopted an official policy covering contributions generated with AI tools and large language models. The move sets clear rules for how such code can be submitted to the open-source project. The announcement is circulating in Linux and open-source communities, where projects are increasingly defining their stance on AI-assisted development.

  28. 28
    PostHog's Jeeves applies reasoning to decision models●Jeeves. Reasoning improves Jev-like decision modelsYhn2323 h ago

    PostHog has released Jeeves, an open-source project on GitHub that uses reasoning to improve Jev-like decision models. The project is drawing attention among developers, ranking highly on Hacker News with over 230 upvotes. Commenters appear interested in how large language model reasoning can be layered onto structured decision-making frameworks for automation.

  29. 29
    Three unpatched critical flaws disclosed in LightLLM▼🚨 LightLLM Mass Disclosure — 3 CVEs, no patch CVE-2026-103040 (CVSS 9.8) — unauthenticated RCE, router profiler RPyC CVEMmastodonTechnologyAI43 h ago

    Three vulnerabilities in LightLLM, an open-source large language model serving framework, have been disclosed without an available patch. The most serious, CVE-2026-103040, is rated 9.8 and allows unauthenticated remote code execution via the router profiler RPyC interface. A similar flaw, CVE-2026-103041, also rated 9.8, affects the embed cache RPyC service, while CVE-2026-103042, rated 7.5, enables memory exhaustion through the NCCL control channel. Security researchers are urging exposed deployments to restrict network access.

  30. 30
    Raschka traces text classification from bag-of-words to LLMs●Language models for text classification: From bag-of-words to JevYhn8913 min ago

    Machine learning researcher Sebastian Raschka has published a deep-dive on the history of text classification, walking through how the field moved from bag-of-words methods such as Naive Bayes and logistic regression to embeddings and modern large language models. The piece examines how much accuracy improved at each stage and what older techniques still offer today, drawing discussion from developers weighing cost against performance.

  31. 31
    Routing LLM traffic across providers with TCP-style congestion control▼Routing LLM traffic across inference providers with TCP-style congestion controlYhnWorldUS Politics753 min ago

    Engineers are discussing a proposal to route large language model inference traffic across multiple providers using congestion control methods borrowed from TCP. The approach dynamically shifts requests toward faster or more reliable providers, similar to how internet protocols manage network congestion. Commenters see it as a practical answer to inconsistent latency and availability across AI inference services.

  32. 32
    TurboGPT trains tiny 22KiB transformer in 13 seconds●Show HN: TurboGPT: train 22KiB transformer in 13sYhnWarMiddle East417 h ago

    A developer known as lostmsu has released TurboGPT, an open-source project on GitHub that trains a compact 22KiB transformer model in roughly 13 seconds. The tool is drawing attention from machine learning enthusiasts interested in fast, lightweight training experiments that can run without large compute budgets.

  33. 33
    Thomson Reuters Unveils Thomson, Its Own Legal AI Model●Meet Thomson: The LLM built by Thomson Reuters✉newsTechnologyAI1 h ago

    Thomson Reuters has introduced Thomson, a large language model it built in-house, aimed at its legal and professional information products. The company positions the model as purpose-built for legal work rather than a general-purpose chatbot. The announcement comes as legal publishers race to develop AI tools that keep pace with competitors in the legal tech space.

  34. 34
    Where Is the Predicted AI Intelligence Explosion?●Where's the "Intelligence Explosion"?Yhn2815 h ago

    Economist Noah Smith asks why the much-discussed 'intelligence explosion' — the idea that AI will rapidly improve itself and trigger runaway progress — has not materialised despite advances in large language models. The essay argues that predictions of sudden, self-accelerating AI capability gains have so far not matched reality, a critique likely to draw debate among AI researchers, economists and tech commentators.

  35. 35
    New CVE Alert Issued for ModelTC LightLLM●CVE Alert: CVE-2026-103042 - ModelTC - LightLLM - https://www. redpacketsecurity.com/cve-aler t-cve-2026-103042-modeltc-MmastodonTechnologyCybersecurity03 h ago

    A security advisory has been published for CVE-2026-103042, a vulnerability affecting LightLLM, the large language model inference server developed by ModelTC. Threat intelligence accounts are circulating the alert to warn organisations running the software to review the flaw and check whether patches or mitigations are available.

  36. 36
    AMD driver update boosts Radeon AI performance up to 23%●🤖 AMD boosting AI/LLM performance for Radeon iGPUs as much as 18~23% with Linux 7.4 submitted by /u/Fcking_Chuck [link]MmastodonTechnologySoftware07 h ago

    AMD is delivering significant AI and large language model performance gains for its Radeon integrated graphics, with improvements of roughly 18 to 23 percent arriving via the Linux 7.4 driver. The gains matter for users running AI workloads on budget and portable systems that rely on integrated GPUs rather than discrete graphics cards. Linux users and AI enthusiasts are discussing what the update means for local LLM performance on AMD hardware.

  37. 37
    Developers Debate Subscriptions Over Rising AI API Costs●Developers Debate Subscriptions Over AI API Costs𝕏xSE5521 h ago

    Developers are weighing whether to switch their apps and services from pay-per-use AI APIs to flat subscription models as API costs for large language models keep climbing. Supporters of subscriptions say predictable pricing protects margins and simplifies billing for users, while critics argue usage-based pricing is fairer and subscriptions can lead to losses when heavy users consume more AI compute than they pay for. The debate has split developer communities, with many sharing cost breakdowns and real-world examples of both approaches.

  38. 38
    New Scientist examines AI's impact on mathematical research●Very interesting article in New Scientist page 5 issue 3613 about the effects of AI and LLM use on # mathsresearch . TheMmastodonTechnologyAI11 h ago

    A New Scientist article in issue 3613 argues that AI and large language models are set to change how progress is made in mathematics. The concern raised is that mathematicians could end up spending much of their time reviewing and filtering AI-generated output rather than doing original work. Readers are debating whether AI tooling will accelerate discovery or burden researchers with checking machine-generated results.

  39. 39
    Benchmark finds AI models inflate security vulnerability severity●Every model (incl. Jev) we tested inflates security finding severityYhnSportFootball89 h ago

    Security firm Casco reports that every large language model it tested, including its own Jev model, inflated the severity of security findings when scoring vulnerabilities, overstating risk compared to expected CVSS ratings. The company published a benchmark detailing the results, prompting discussion about how far AI-generated severity scores can be trusted in security workflows.

  40. 40
    Amidi Brothers Release Illustrated Guide to Transformers and LLMs●Super Study Guide: Transformers & Large Language Models by Afshine Amidi and Shervine Amidi 📖 on Leanpub! A clear, illusMmastodonTechnologyAI111 h ago

    Afshine Amidi and Shervine Amidi have published 'Super Study Guide: Transformers & Large Language Models' on Leanpub. The book offers a clear, illustrated introduction to large language models, covering key concepts and practical applications. The authors say it is suited to work projects, interview preparation, or personal learning, adding to their well-known series of study guides on machine learning topics.

Repos