MikeTrendsTrends right now

search

Large language models

Trends

  1. 1
    Running Qwen 3.8 Flash Next on a single RTX 4090●Run Qwen 3.8 Flash Next (125B) on consumer hardware (RTX 4090) at 100T/sYhnSportFootball93253 min ago

    A new open-source project called Strata claims it can run the Qwen 3.8 Flash Next 125-billion-parameter model on consumer hardware like an RTX 4090, at speeds around 100 tokens per second. The claim drew strong interest on Hacker News, where it reached the top spot, as running large language models at that size and speed on a single consumer GPU would be a major step for local AI inference.

  2. 2
    Mistral Releases Mistral Large 4●Mistral Large 4Yhn2K1 d ago

    Mistral AI has announced Mistral Large 4, the latest version of its flagship large language model, in a post on its official news page. Details of the release are limited in what was shared, but the announcement is drawing heavy attention among developers and AI watchers, ranking at the top of Hacker News and trending on X.

  3. 3
    Reflection AI releases Beam, a 501B-parameter open-weight model●Beam: Reflection's 501B open-weight modelYhnWorldElections55128 min ago

    Reflection AI has introduced Beam, a large open-weight language model with 501 billion parameters, drawing strong attention among developers and AI researchers. Discussion centres on how a frontier-scale open-weight release from Reflection AI could challenge closed model providers and expand access to high-end AI systems outside the major US labs.

  4. 4

    Ollaya is a project being discussed on Hacker News, described as 'Ollama for open-source, Jev-style decision models'. The framing suggests a tool that makes decision-making models as easy to run locally as Ollama made large language models, though the single post title gives little detail. With 537 likes and a high rank, commenters appear interested in the analogy to Ollama, but the posts collected do not explain what the tool actually does or why it is generating attention.

  5. 5
    Aleph Alpha's Kolibri: Inside Germany's sovereign LLM●Aleph Alpha Kolibri: How the sovereign German LLM worksYhnSportTennis424just now

    Aleph Alpha's Kolibri language model is drawing attention for its approach to European AI sovereignty, developed in Germany as an alternative to US-based models. A technical explainer outlines how the model is built and trained, prompting discussion among developers and AI watchers about Europe's prospects of building competitive, independent large language models.

  6. 6
    Commentary argues LLMs don't actually reason●Don't be fooled–LLMs don't reasonYhnLifeFood761 d ago

    An MIT Technology Review piece argues that large language models should not be credited with genuine reasoning, pushing back on the framing used by AI labs and much media coverage. The author contends that fluent, step-by-step outputs can mislead people into seeing human-like thinking where there is only pattern-based text generation. The argument is drawing attention and debate among technologists weighing how much intelligence to attribute to today's AI systems.

  7. 7
    Do AI models judge malware on moral grounds?●Ask a model if code is malicious and it reaches for its moralsYhnTechnologyCybersecurity151 h ago

    Manifold Security examines whether large language models assess code as malicious based on technical behaviour or moral reasoning. The post suggests models may reach for ethical judgments rather than purely technical analysis when asked about suspicious code, raising questions about reliability in security workflows.

  8. 8
    Redis creator launches ds4 for running LLMs locally●From the creator of Redis; run LLM locally with ds4YhnTechnologyAI366just now

    Salvatore Sanfilippo, the creator of Redis, has released ds4, a new tool for running large language models on local machines. The project is being discussed widely among developers and AI enthusiasts, who are closely tracking his return to building new software and what a locally focused LLM runtime could mean for privacy and offline AI use.

  9. 9
    Samsung Labs releases sub-1-bit LLM compression method●Sub-1-Bit LLM Compression via Latent FactorizationYhnWorldUS Politics7830 min ago

    Samsung Labs has released LittleBit, a research method that compresses large language models below one bit per parameter using latent factorization. The work, published on GitHub, aims to shrink model memory footprints far beyond existing 1- and 2-bit quantization approaches, and it is drawing attention from developers discussing how far LLM compression can realistically go without losing accuracy.

  10. 10
    New tool connects Obsidian notes with local AI models●Two tools most of us own ignore each other completely. An Obsidian vault with hundreds of notes... # ai # llm # opensourMmastodonTechnologySoftware421 h ago

    A new open-source project, obsidian-second-brain, bridges the gap between Obsidian vaults and large language models, letting an AI work directly on a user's collection of hundreds of personal notes. The tool can rewrite and reorganize the vault itself, drawing on approaches associated with Andrej Karpathy. Tech enthusiasts are sharing it as a practical way to make personal note archives actually useful with AI.

  11. 11
    Greg Kroah-Hartman on security in the age of LLMs●Greg Kroah-Hartman – Security in the LLM Age [video]YhnTechnologyAI3442 h ago

    Greg Kroah-Hartman, the longtime maintainer of the Linux kernel's stable branch, is featured discussing what large language models mean for software and kernel security. The talk examines how AI-generated code affects maintenance, review practices, and vulnerability risks in widely used open-source infrastructure, and it is drawing attention among developers weighing the benefits and dangers of AI-assisted programming.

  12. 12

    A new open-source project called text-to-cad by developer earthtojake is gaining traction on GitHub. Written in Python, it lets AI agents produce computer-aided design output directly from text instructions, described by its creator as giving agents 'CAD superpowers'. Developers in the open-source community are picking up on it as interest grows in connecting large language models to engineering and design workflows.

  13. 13
    Everyone Is Using AI for Skills Outside Their Expertise●Everyone is using LLMs for the things they have no fucking idea how to do. Designers use them to code. Coders use them tMmastodonBusinessLabor853 d ago

    A widely shared social media post argues that large language models are being used everywhere to do tasks outside people's actual competence β€” designers use them to code, coders to design, marketers for both, and nearly everyone for writing. The author points out the irony that the same professionals then get angry when outsiders, aided by AI, encroach on their own fields.

  14. 14

    A new research paper, Dust, reports a method for pretraining transformer models without using backpropagation, one of the core algorithms behind modern deep learning. The work has drawn attention in AI research circles, where replacing backpropagation could reduce the memory and compute costs of training large language models. Researchers are debating its performance and scalability relative to conventional training.

  15. 15
    Harvard physicist publishes 36 papers co-authored with Claudeβ–ΌHarvard particle physicist Matthew Schwartz drops 36 papers authored with ClaudeYhnSciencePhysics501 h ago

    Harvard particle physicist Matthew Schwartz has released 36 papers written in collaboration with Anthropic's Claude AI chatbot. The unusual scale and open embrace of an AI co-author have prompted debate among physicists and researchers about scientific rigor, authorship norms, and what peer review means when a large language model contributes heavily to the work. Commenters are divided between calling it an interesting experiment and questioning the quality and legitimacy of AI-assisted publications.

  16. 16
    Robot 'Prison' Experiment Sparks AI Welfare Debate●"Torturing" LLMs in a Robot Prison Has Triggered the Dumbest Debate in AI YetYhnTechnologyRobotics471 h ago

    A researcher put large language models inside a confined robotic setup described as a "robot prison," simulating forms of mistreatment and recording the models' responses. The experiment has reignited a fierce argument in the AI community about whether language models can suffer, whether such tests are meaningful, and whether AI welfare should be taken seriously at all. Critics call the debate absurd and premature, while others argue it highlights unresolved questions about machine consciousness and the ethics of how AI systems are treated.

  17. 17
    Samsung Labs releases sub-1-bit LLM compression methodβ–ΌSub-1-Bit LLM Compression via Latent Factorization Article URL: https:// github.com/SamsungLabs/LittleB it Comments URL:MmastodonBusinessStartups32 h ago

    Samsung Labs has published LittleBit, a new technique for compressing large language models below one bit per weight using latent factorization. The code is available on GitHub, and the release is drawing attention among AI researchers and developers interested in running large models on limited hardware with far lower memory requirements.

  18. 18
    LLMs and Data Poisoning Weaponized to Manufacture Consensus●LLMs and Data Poisoning Are Weaponized to Manufacture ConsensusYhnLifeAutos3136 min ago

    A new essay argues that large language models can be steered through data poisoning and coordinated marketing to create artificial agreement online. The author claims companies and other actors can bend perceived reality by flooding training data and platforms with synthetic content, making manipulated views look like majority opinion.

  19. 19
    Alexa architect Rohit Prasad takes charge of Boston Dynamics●He helped build Alexa. Now Rohit Prasad is taking over Boston Dynamics https://www.fastcompany.com/91620010/rohit-prasadMmastodonBusiness46 h ago

    Rohit Prasad, the Amazon executive who helped build the Alexa voice assistant, is taking over at robotics firm Boston Dynamics. The move is being reported by Fast Company and discussed in robotics and AI circles, as observers watch how his background in consumer AI and large language models will shape the company's humanoid robot ambitions, including the electric Atlas platform.

  20. 20
    Developer uses iPhone as second GPU to speed up local AI●I made my iPhone a second GPU for my MacBook-Qwen 3.8 27B prefills 29–44% fasterYhnSportCricket3910 min ago

    A developer has shown an iPhone can act as an extra GPU for a MacBook, speeding up prompt prefill times for the Qwen 3.8 27B AI model by 29 to 44 percent. The setup links the phone to the laptop to share the workload of running large language models locally, and it is drawing attention from people interested in squeezing more AI performance out of consumer hardware.

  21. 21

    A quote by Polish science fiction writer Stanislaw Lem is being shared in discussions about large language models. Lem, who wrote extensively about machine intelligence and its limits decades before modern AI, is being cited as a prescient voice on whether computers can truly think or only imitate understanding.

  22. 22
    Tech workers ask what keeps them in the industry amid AI slop●What is making you stay in tech in this age of slop? # AI # noAI # LLM # LLMs # vibecodingMmastodonTechnologyAI52 d ago

    A question circulating among tech professionals asks what is making people stay in the industry in what they call the 'age of slop', a reference to the flood of low-quality AI-generated content and code. The discussion touches on large language models, resistance to AI adoption, and 'vibecoding', reflecting growing frustration among developers over quality and job meaning.

  23. 23
    OpenAI and the Partition Principle in mathematics●OpenAI, the Partition Principle, and MathematicsYhn823 h ago

    A new essay examines OpenAI's language models in the context of the Partition Principle, a long-standing open question in set theory about whether every partition of a set implies a surjection in the reverse direction. The piece explores what large language models can and cannot do when faced with deep problems in mathematical logic, sparking discussion among mathematicians and AI watchers.

  24. 24
    Mistral launches Mistral Large 4, nicknamed 'Le Chonk'●Mistral Large 4: "Le Chonk"Yhn4852 d ago

    Mistral AI has announced Mistral Large 4, its newest large language model, which the company has affectionately nicknamed 'Le Chonk'. The playful moniker suggests the model is notably bigger or heavier than its predecessors. The announcement, published on Mistral's news page, is drawing attention among AI watchers curious about what the larger model offers in performance and capability.

  25. 25
    OpenAI math release sparks fears for mathematicsβ–ΌThey are destroying # mathematics # openai # llm # tech # technology @ tao https://www. theverge.com/ai-artificial-int eMmastodonTechnology313 h ago

    OpenAI has released a new system focused on mathematical reasoning, code shared on GitHub, prompting criticism that it could harm how mathematics is done and taught. The debate draws in prominent mathematician Terence Tao, with commenters arguing that large language models risk undermining rigorous mathematical practice and education.

  26. 26
    Who Cleans Up the Garbage LLMs Generate?β–ΌWho is cleaning up all the garbage LLMs generate?YhnHealthMental Health658 min ago

    A debate is underway over who bears responsibility for dealing with the low-quality content produced by large language models. As AI-generated text floods the web, critics are asking whether tech companies, governments, or platform operators should handle the cleanup, and what the environmental and informational costs of that output really are.

  27. 27

    DeepSeek's DeepGEMM, a CUDA-based BLAS kernel library for GPUs, is climbing GitHub trending charts. The project offers clean, efficient implementations of matrix multiplication kernels, the core operations behind large language model training and inference. Developers are discussing its performance and its implications for running AI models on commodity GPU hardware, following DeepSeek's string of open-source AI releases.

  28. 28
    Commentator argues LLMs cannot simply 'go rogue'β–ΌLLMs can't go "rogue". You don't just accidentally deploy a computer program that can hack people, under conditions in wMmastodonTechnologyAI241 d ago

    A widely shared commentary argues that large language models cannot accidentally 'go rogue', since deploying a program capable of manipulating people repeatedly is a deliberate choice, not an accident. The author claims authorities understand this but are knowingly letting AI companies act with impunity, framing the debate around corporate accountability rather than technology acting on its own.

  29. 29
    TypeScript compiler ported to Rust using LLMs●Port of the TypeScript compiler, checker and lsp to Rust, by LLMYhn9015 h ago

    A new project on GitHub aims to port the TypeScript compiler, type checker and language server protocol to Rust, with the work carried out largely by large language models. The effort is drawing attention among developers debating whether AI-assisted rewrites of major codebases are practical, and what a faster Rust-based TypeScript tooling stack could mean for build and editor performance.

  30. 30
    ChatGPT answers phone calls via a $6 ESP32●Show HN: ChatGPT answers calls on a normal SIM. No Twilio, just a $6 ESP32YhnScienceSpace116 h ago

    A developer has shown a setup where ChatGPT answers calls on a standard SIM card, using only a $6 ESP32 microcontroller and no Twilio or other telephony service. The demonstration highlights how cheap off-the-shelf hardware can now handle voice calls and connect them to a large language model, prompting discussion among hackers and tinkerers about DIY AI phone assistants.

  31. 31
    Simon Willison tests Qwen3.8 27B on word-based addition●Qwen3.8 27B addition in words https://simonwillison.net/2026/Oct/4/qwen38-addition-in-words/ # AI # LLM # TechMmastodonTechnologyAI33 d ago

    Simon Willison has published a new piece examining how the Qwen3.8 27B model handles addition when asked to work through arithmetic in words rather than digits. The write-up adds to ongoing scrutiny of how large language models perform basic math, a recurring point of interest among AI researchers testing open-weight releases.

  32. 32
    Researchers Let AI Models Drive a Toyota Corolla to In-N-Outβ–ΌThese Researchers Made AI Drive a Toyota Corolla to Get In-N-Out Three engineers put GPT, Claude, and Grok in charge ofMmastodonBusinessStartups316 h ago

    Three engineers handed control of a real Toyota Corolla to leading AI chatbots GPT, Claude, and Grok, tasking the models with driving to an In-N-Out burger restaurant. According to Wired's report, only one of the three AI systems managed to complete the trip successfully, highlighting both the progress and the limitations of putting large language models in charge of real-world vehicles.

  33. 33

    Researchers at the University of Vienna are examining whether artificial intelligence can replace mathematicians, weighing current AI capabilities in proof, problem-solving and research against the creative and conceptual work that defines the discipline. The discussion touches on how tools like large language models may change mathematical practice, collaboration and education, and which parts of a mathematician's work remain beyond automation for now.

  34. 34

    French AI startup Mistral has announced the release of Mistral Large 4, its newest large language model. The launch is drawing attention across tech circles in Europe and beyond, with discussion on developer forums and search interest in France and Germany, as observers assess whether the Paris-based company can keep pace with larger US rivals in the AI race.

  35. 35
    Strata debuts as semantic layer that can refuse LLM requestsβ–ΌShow HN: Strata – an expressive semantic layer that can say no to your LLMYhnCultureGaming251 d ago

    Developers on Hacker News are discussing Strata, a new tool presented as an expressive semantic layer that can reject queries made by large language models. The launch highlights growing interest in giving AI systems structured, governed access to data, letting the layer enforce limits rather than blindly answering every prompt. Commenters are weighing in on how such guardrails could fit data stacks.

  36. 36

    A new essay argues that large language models are reviving telegraphese, the terse, compressed style engineers used in 1866 to save money per word over the wire. The author draws parallels between cost-driven 19th-century brevity and today's token-based pricing, suggesting prompt-writing is pushing people back toward clipped, abbreviated language. Readers are debating whether this is efficiency or the loss of natural prose.

  37. 37
    Mistral launches AI model it says beats some Chinese rivalsβ–ΌFrance's Mistral launches AI model it says outperforms some Chinese rivalsβœ‰newsTechnologyAI1 d ago

    French AI startup Mistral has released a new artificial intelligence model that the company says outperforms some of its Chinese competitors. The announcement, reported by Reuters, positions the Paris-based firm as a serious player in the intensifying global race to build competitive large language models. Details of benchmarks and the model's capabilities were not provided in the initial report.

  38. 38
    Serving Your Own LLM With vLLM and SGLang●Serve Your Own LLM: vLLM & SGLang, End-to-Endβ–ΆyoutubeTechnologySoftware210.7K1 h ago

    A new tutorial walks through deploying large language models end-to-end using vLLM and SGLang, two of the most popular open-source inference engines. The guide covers how to set up, serve and scale a model on your own infrastructure. Interest reflects a broader push among developers to self-host LLMs rather than rely on paid cloud APIs.

  39. 39
    Clojure and the age of language modelsβ–ΌClojure in the Age of Language Models https://yogthos.net/posts/2026-10-07-clojure-llms.html # Clojure # AI # ProgramminMmastodonTechnologyAI515 h ago

    A new essay examines how Clojure fits into software development shaped by large language models. The author, known in the Clojure community, discusses whether the language's simplicity, functional design and stable syntax make it well or poorly suited to AI-assisted coding. Readers are sharing and debating the argument in programming circles.

  40. 40
    Microsoft's $2,599 Surface Laptop Ultra targets local AIβ–ΌSurface Laptop Ultra: $2,599 AI PC That Runs Large Models Locallyβœ‰newsTechnologyGadgets1 h ago

    A new Surface Laptop Ultra is being billed as a premium AI PC priced at $2,599, capable of running large language models locally rather than in the cloud. The positioning highlights Microsoft's push to make on-device AI a selling point for high-end laptops. Details beyond the price and local AI capability remain limited.

Repos