MikeTrendsTrends right now

search

Large language models

Trends

  1. 1
    Samsung Researchers Unveil Sub-1-Bit LLM Compression Method●Sub-1-Bit LLM Compression via Latent FactorizationYhnWorldUS Politics8822 min ago

    Samsung Labs has released LittleBit, a new technique for compressing large language models to less than one bit per weight using latent factorization. The work, published on GitHub, promises to shrink memory footprints of big models far beyond existing quantization methods, making it possible to run them on modest hardware. Developers and AI researchers are discussing its potential to cut deployment costs and bring large models to consumer devices.

  2. 2
    AirLLM runs 70B models on a 4GB GPU●lyogavin/airllm⬢github54111 min ago

    The open-source project AirLLM, hosted on GitHub by developer lyogavin, enables inference of 70-billion-parameter large language models on a single GPU with only 4GB of memory. The approach dramatically lowers the hardware barrier for running large AI models locally, and it is drawing attention from developers interested in accessible LLM deployment.

  3. 3

    TensorFlow, Google's open-source machine learning framework, is trending on GitHub this week. The repository provides tools for building and training machine learning models, and is written largely in C++ with interfaces for Python and other languages. The posts visible are simply links to the repository with its standard description, so there is no specific release, announcement, or discussion evident from the snippets. Trending likely reflects renewed attention from developers, but the exact trigger is not clear from the posts.

  4. 4
    Debate rages over developers claiming 'no AI' after using AI tools●This fuckin "I made this myself with No A.I .. but then after release I used LLMs and AI to bug fix and create new featuMmastodonTechnologySoftware51 h ago

    An argument is spreading online over game and software developers who market their work as made entirely without AI, only to later admit they used large language models for bug fixes and new features after release. Critics call this misleading wording and say developers should be upfront about how AI was used. Supporters of the 'no AI' label argue buyers deserve accurate disclosure about how products are actually made.

  5. 5
    Connecting LLM reasoning to deterministic financial engineering with MCP▼A deep dive into bridging LLM reasoning with deterministic financial engineering using MCP connectors. # ai # mcp # finaMmastodonBusinessFinance33 min ago

    A new technical deep dive explores how large language model reasoning can be combined with deterministic financial engineering through MCP connectors. The approach aims to let AI models handle flexible reasoning while delegating high-precision financial calculations to reliable, rule-based systems. The piece is being shared among developers and finance technologists interested in AI integration and software architecture.

  6. 6
    Can ChatGPT beat the stock market? AI trading tested●Can ChatGPT beat the stock market? AI stock trading now goes beyond research: LLM agents can analyze news, portfolios anMmastodonBusiness44 min ago

    AI stock trading is moving beyond research assistance: large language model agents can now analyze financial news, review portfolios and connect directly to brokerages. Studies show real predictive signals in financial text, but there is no proof yet of a reliable market-beating machine. Observers are weighing genuine analytical gains against hype and the risk of overtrusting automated agents with money.

  7. 7
    AI Data Poisoning Used to Manufacture False Consensus●LLMs and Data Poisoning Are Weaponized to Manufacture ConsensusYhnLifeAutos3128 min ago

    A new essay argues that large language models and data poisoning techniques are being deliberately used to manipulate public opinion and manufacture artificial consensus. Drawing on examples from marketing and political influence, the author warns that synthetic content can bend perceived reality at scale, making coordinated manipulation cheaper and harder to detect than ever before.

  8. 8
    Pennsylvania credit union to offer Google Gemini chatbot to members●PSECU, a credit union in Harrisburg, Pennsylvania, to offer Google Gemini-based bot to members Consumers have started tuMmastodonBusinessBanking21 h ago

    PSECU, a credit union based in Harrisburg, Pennsylvania, is set to offer its members an assistant built on Google's Gemini technology. The move reflects a broader shift in consumer finance, as customers increasingly turn to large language models such as ChatGPT, Claude and Gemini for financial guidance. The story was reported by American Banker, highlighting how smaller financial institutions are adopting generative AI tools to keep pace with member expectations.

  9. 9
    Seven AI models asked to estimate Trump's IQ▼I Let 7 AIs Estimate Trump's IQ. And It's Exactly What You'd Expect▶youtubeTechnologyAI2.3M1 h ago

    A widely shared experiment asked seven different AI chatbots to estimate Donald Trump's IQ, and the results are being presented as predictable. The item is drawing large engagement as viewers compare the models' answers and debate whether AI estimates of public figures' intelligence are meaningful at all. Reactions split between amusement at the answers and scepticism that language models can credibly assess IQ without any testing.

  10. 10
    Commenters push back on the idea that AI adoption is inevitable●There's little that's "inevitable" about AIYhnCultureTheatre4032 min ago

    A new essay argues that large language models are not an unavoidable force, challenging the widespread narrative that AI's dominance is certain. It suggests adoption depends on choices by companies, regulators and users rather than technological destiny. The argument is drawing attention online, with readers debating whether AI hype has outpaced its practical value and whether the 'inevitability' framing serves industry interests more than reality.

  11. 11
    AI models weigh morality when judging malicious code●Ask a model if code is malicious and it reaches for its moralsYhnTechnologyCybersecurity152 h ago

    Manifold Security published an examination of how large language models respond when asked whether code is malicious, finding that models tend to introduce moral reasoning into their assessments rather than giving purely technical answers. The piece has drawn attention on Hacker News, where readers are debating whether moral framing helps or hinders reliable malware analysis and what it reveals about how these models were trained.

  12. 12
    Who cleans up the garbage LLMs generate?●Who is cleaning up all the garbage LLMs generate?YhnHealthMental Health650 min ago

    A debate is growing over who is responsible for cleaning up the low-quality, repetitive, or misleading content that large language models produce at scale. Critics argue that AI-generated text is flooding the internet, polluting search results and online communities, and that no one — not model developers, publishers, or platforms — has taken clear ownership of the cleanup. Others counter that filtering tools and human moderation are already adapting, but agree the volume of synthetic content is rising faster than the systems meant to contain it.

  13. 13
    Mount Sinai Study Finds Safety Prompts Improve AI Clinical Decisions▼Mount Sinai Study Finds Safety Prompts Can Help AI Models Make Safer Clinical Choices✉newsTechnologyAI1 h ago

    Researchers at Mount Sinai report that adding safety prompts to AI models can guide them toward safer clinical recommendations. The study suggests that simple instruction-based safeguards, rather than costly retraining, may reduce risky medical advice from large language models. The findings are drawing attention as hospitals weigh how to deploy AI tools in patient care settings.

  14. 14
    Egocentric video seen as key to robot AI brains●“In the past few years, researchers have found that first-person or egocentric video is a crucial ingredient for the # AMmastodonTechnologyAI21 h ago

    Researchers report that first-person, egocentric video has become a crucial ingredient for training the AI models that act as brains for robots. Much as large language models absorb vast amounts of human writing to learn language, robots appear to benefit from visual data captured from a human point of view, helping them understand how tasks are performed. The observation highlights the growing overlap between AI research in chatbots and embodied robotics.

  15. 15
    New Self-Pruning Transformer Promises Extreme KV-Cache Compression●A Self-Pruning Transformer: Extreme KV-Cache Compression w/Universal AttentionYhnWorldUS Politics61 h ago

    A new paper on arXiv introduces a self-pruning transformer architecture claiming extreme KV-cache compression using a universal attention mechanism. KV-cache memory is a major bottleneck for running large language models efficiently, so techniques that shrink it could cut inference costs and enable longer contexts on ordinary hardware. Early reactions are focused on the technical details and whether the compression claims hold up in practice.

  16. 16
    Why dexterity is physical AI's real bottleneck▼Why dexterity is physical AI’s real bottleneck✉newsTechnologyRobotics2 h ago

    Industry discussion is focusing on robotic dexterity as the main obstacle holding back physical AI. While large AI models have advanced rapidly in language and vision, robots still struggle with fine motor tasks like grasping and manipulating unfamiliar objects, which limits deployment in factories, warehouses and homes. Analysts argue that solving manipulation, not intelligence alone, is the key challenge for humanoid and industrial robotics going forward.

  17. 17

    The Hugging Face Transformers library, an open-source Python framework for building and running state-of-the-art machine learning models across text, vision, audio, and multimodal tasks, is attracting significant attention from developers worldwide. The project supports both inference and training, making it a core tool for AI engineers and researchers. Its popularity reflects the continued surge in demand for accessible machine learning infrastructure as adoption of large language and multimodal models accelerates.

  18. 18
    Developers joke about burnout and escaping to gardening●then they ask why we burn out and go gardening # dev # developer # c # assembly # coding # programming # programmer # opMmastodonTechnologySoftware41 h ago

    Programmers are sharing a joke about burnout in software development, suggesting many are so exhausted by coding work that they dream of quitting to grow gardens instead. The quip touches on familiar pressures in the industry, from low-level languages like C and assembly to modern JavaScript, TypeScript and the growing demands of AI and large language models, resonating with developers who feel overworked.

  19. 19
    Book Publishers Quietly Turn to AI, Staff Push Back●Book Publishers Are Quietly Using More AI. Staff Are Revolting WIRED Workers at three major publishing houses tell WIREDMmastodonCultureBooks102 h ago

    Workers at three major publishing houses have told WIRED that publishers are increasingly using large language models and generative AI for tasks including publicity materials, cover art and back cover copy, often without telling authors or readers. Staff inside the industry are pushing back against the practice, raising concerns about quality, transparency and the erosion of creative jobs. The report has reignited debate over how much AI belongs in book publishing.

  20. 20
    Deep-Dive Walkthrough of LLM Application Development Published▼LLM Application Development A deep-dive walkthrough of building applications on top of large language models — coveringMmastodonTechnologySoftware32 h ago

    A new guide walks developers through building applications on top of large language models. It explains how LLMs actually work and what that implies for application design, prompt engineering, and the system, user and assistant message roles. It also covers few-shot prompting and structured output techniques.

  21. 21
    Will AI sour people on social media?●💬 Wird KI (LLMs) die Lust auf # SocialMedia vermiesen und die Zahl der Nutzer.innen sinken?MmastodonTechnologyInternet72 h ago

    A discussion is under way about whether large language models and AI-generated content could ruin the appeal of social media and drive user numbers down. The concern is that feeds increasingly filled with machine-written posts and synthetic replies may erode trust and enjoyment, pushing people to spend less time on the platforms or leave them altogether.

  22. 22
    Developer sparks debate over AI and labor exploitation●RE: https:// mastodon.gamedev.place/@aeva/1 17413312861881267 "AI" (LLMs) are just a way for the rich to have their cakeMmastodonTechnologyAI72 h ago

    A game developer's commentary arguing that large language models amount to a way for the wealthy to have their cake and eat it is circulating among tech workers. The argument holds that if AI systems are merely statistical models built on human labor, they constitute copyright infringement, while claims of sentience are used to dodge accountability. Responses are echoing the critique of how AI companies treat creative and intellectual work.

  23. 23
    Neurosymbolic AI pitched as path beyond scaling up language models●# Neurosymbolic AI – Why the Future of # ArtificialIntelligence Needs More Than Just # Data 🤖 🧠 🔮 Simply # scaling up maMmastodonTechnologyAI13 h ago

    Supporters of neurosymbolic AI argue that simply scaling up large language models and feeding them ever more data will not deliver true artificial general intelligence. They contend the field needs architectures combining neural networks with symbolic reasoning to move beyond pattern matching, and the debate over scaling versus hybrid approaches is drawing renewed attention among AI researchers and commentators.

  24. 24
    CrowdStrike says lone hacker used AI tools to breach South Korean banks●🤖 CrowdStrike: a single operator used ARTEX, an open-source Chinese agentic pentest tool, plus LLMs, to hit South KoreanMmastodonTechnologyCybersecurity15 h ago

    CrowdStrike reports that a single operator breached South Korean banks between late September and early October using ARTEX, an open-source Chinese agentic penetration testing tool, alongside large language models. Roughly 68,000 records were exposed through a loan-inquiry service and an employee mobile portal. Analysts note the attackers relied on conventional flaws and cheap, off-the-shelf offensive AI tooling rather than sophisticated custom malware.

  25. 25
    Anthropic launches Claude Haiku 5.5 as Mistral releases Mistral Large 4●Anthropicが小型AIモデル「Claude Haiku 5.5」を発表/Mistral AIが大規模言語モデル「Mistral Large 4」を公開(ITmedia PC USER) https://www. yayafa.com/MmastodonTechnologyAI15 h ago

    Anthropic has announced Claude Haiku 5.5, a small, lightweight AI model, while French startup Mistral AI has released its new large language model Mistral Large 4. The two announcements, reported by ITmedia PC USER, came in quick succession and are being discussed among AI watchers as the latest sign of intensifying competition in both compact and frontier-scale models.

Repos