MikeTrendsTrends right now

search

local AI models

Trends

  1. 1
    Open-Source Edge Inference Engine Runs Large AI Models on Robots 10.7x Faster▼10.7x Faster: This Open-Source Edge-Side Inference Engine Enables Robot Bodies to Run Large Models Without Lag✉newsTechnologySoftware4 d ago

    A new open-source edge-side inference engine claims a 10.7x speedup, allowing robot hardware to run large AI models locally without lag. The technology targets real-time on-device inference for robotics, reducing reliance on cloud computing. Discussion is centered on its performance gains and what faster local inference could mean for embodied AI and robot deployments.

  2. 2
    Commenters argue for public alternatives to corporate AI control●"Turns out there are more options than “hand it to corporations” and “throw every GPU into the sea.” Who knew. Public inMmastodonTechnologyAI56 d ago

    A widely shared commentary argues that debates over artificial intelligence wrongly frame the choice as either corporate control or abandoning the technology entirely. It lists alternatives: public infrastructure, worker co-operatives, open-weight models, union bargaining, regulation, shorter work weeks, local models, shared gains and human oversight, while conceding the details are not fully worked out.

  3. 3
    AI debate: on-device compute or data centers?●🤖 Will the AI compute crunch be solved on-device or in data centers? I build iOS apps and I'm pushing as much as possiblMmastodonTechnologyAI25 d ago

    An iOS developer is weighing whether the growing demand for AI computing power will ultimately be met on devices or in data centers, saying they push as much processing on-device as possible for privacy and cost reasons. They note Apple is betting on on-device AI, but argue frontier models keep getting bigger, and are asking where others think the balance will land.

  4. 4
    Redis creator launches local LLM tool ds4●From the creator of Redis; run LLM locally with ds4YhnTechnologyAI34549 min ago

    A new tool called ds4, promoted as coming from the creator of Redis, lets users run large language models locally on their own machines. The project, hosted at dwarfstar.sh, is drawing attention in developer circles, with discussion centered on what the Redis creator's involvement means for the credibility and future of local AI tooling.

  5. 5
    Local AI Models Now Run Smoothly on Consumer Gaming Hardware●Local AI Models Run Smoothly on Gaming PCs and Laptops𝕏xSE27614 h ago

    Locally run AI models are reportedly operating smoothly on ordinary gaming PCs and laptops, without cloud servers or subscriptions. The discussion centers on how modern GPUs and increasing memory in consumer machines are enough to handle open-source language models at home. Commenters highlight growing interest in private, offline AI use and note that hardware once bought mainly for games is now doubling as a capable local AI workstation.

  6. 6
    Developer uses iPhone as second GPU to speed up MacBook AI workloads●I made my iPhone a second GPU for my MacBook-Qwen 3.8 27B prefills 29–44% fasterYhnSportCricket2444 min ago

    A developer has shown a setup that turns an iPhone into a secondary GPU for a MacBook, reporting that prefill times for running the Qwen 3.8 27B language model are 29–44% faster. The workaround exploits the iPhone's neural hardware alongside the laptop, drawing interest from hobbyists running large local AI models on Apple hardware.

  7. 7
    180B-parameter LLM runs locally on a laptop without a GPU●GPU 없이 소비자용 노트북에서 180억 파라미터 LLM을 구동하는 POCKET-Darwin-180B. 4비트 GGUF 양자화로 360GB→111GB 압축, 약 $1,400 하드웨어로 로컬 추론 가능. # ai #MmastodonTechnologyAI31 d ago

    A project called POCKET-Darwin-180B is drawing attention for running a 180-billion-parameter language model on consumer hardware with no discrete GPU. Using 4-bit GGUF quantization, the model is compressed from roughly 360GB down to 111GB, enabling local inference on hardware costing about $1,400. Commenters in AI and open-source circles are highlighting it as a sign that frontier-scale models may soon run off the cloud.

  8. 8
    Developer launches pretrained classifiers that run without GPU●Show HN: Local pretrained classifiers, GPU not neededYhnWorldElections81 h ago

    A developer has released Jeffy, an open-source tool offering locally running pretrained image classifiers that do not require a GPU. The project, shared on GitHub and introduced on Hacker News, makes machine learning classification accessible on ordinary hardware. Early response is small but positive, with users showing interest in lightweight, privacy-friendly local inference options.

  9. 9
    Developers Turn to Mac Minis for Running AI Models●Why Developers Are Running AI Models on Mac Minis Instead of Nvidia GPUs✉newsBusinessStartups6 d ago

    Developers are increasingly running AI models on Apple's Mac Mini instead of relying on Nvidia GPUs, according to a Fortune report. The shift is being attributed to the Mac Mini's lower cost and power efficiency, with Apple silicon offering competitive performance for local AI workloads. The trend highlights a challenge to Nvidia's dominance in AI hardware as smaller teams look for cheaper ways to build and test AI applications.

  10. 10
    Codex plugins can now be used inside Pi coding agent●Show HN: Use all Codex Plugins inside Pi I just realized that codex now exposes local server endpoints for all plugins wMmastodonBusinessStartups319 h ago

    A developer has discovered that Codex exposes local server endpoints for all of its plugins without extra authentication, meaning those plugins can be used from any other model or agent harness. A new Pi install package lets users connect to all Codex plugins with a single auth setup. Developer communities are discussing what this means for interoperability between AI coding tools and whether open local endpoints could raise security questions.

  11. 11
    Multi-Token Prediction Boosts RTX 3090 LLM Speed▼Originally published on my blog. Enabling MTP on this RTX 3090 raised generation throughput from... # ai # llm # programMmastodonTechnologySoftware51 d ago

    A developer reports enabling multi-token prediction (MTP) on an RTX 3090 graphics card raised local LLM generation throughput, while questioning whether the speedup affects coding quality. The write-up, originally published on a personal blog, has drawn attention from AI and open-source software communities interested in getting more performance from consumer GPUs for running large language models locally.

  12. 12

    Morocco has released its first open-source AI tools for Darija, the Moroccan Arabic dialect, developed in partnership with French AI firm Mistral. The release marks a step toward building AI systems that understand and serve local languages, which are often poorly represented in mainstream AI models trained mostly on English and standard Arabic.

  13. 13
    Writer swaps Grammarly for a local LLM to keep text private●I replaced Grammarly with a local LLM, and none of my writing leaves my laptop anymore✉newsTechnologyAI49 min ago

    A writer describes replacing Grammarly with a locally running large language model for grammar and writing help, meaning drafts no longer get uploaded to cloud servers. The account highlights growing interest in privacy-first AI tools that run on personal hardware, as users grow uneasy about sending sensitive text to third-party services for editing suggestions.

  14. 14
    Open source tool lets you run Jev locally▼Open source tool distills Jev so you can run it locally✉newsTechnologySoftware3 d ago

    The Register reports on a new open source tool that distills Jev, making it possible to run it on local hardware rather than in the cloud. Distillation shrinks a model so it can run on ordinary machines, lowering cost and keeping data private. The piece describes the tool and what it means for developers wanting offline use.

  15. 15
    Philadelphia Inquirer launches AI tool Scrape for hyperlocal news●The Philadelphia Inquirer built Scrape, an AI tool to surface hyperlocal news https://www.lenfestinstitute.org/solutionsMmastodonTechnology31 d ago

    The Philadelphia Inquirer has built Scrape, an artificial intelligence tool designed to surface hyperlocal news that might otherwise go unreported. The project is highlighted by the Lenfest Institute, which supports the paper and promotes it as a model for local journalism. Observers in tech circles are discussing whether AI can help struggling local outlets cover neighborhood-level stories at scale.

  16. 16
    Telegram bot reads bills locally with Gemma and nagging reminders●A Telegram bot that reads bills with local Gemma and keeps reminding until you pay, with durable reminders on Temporal aMmastodonTechnologySoftware27 h ago

    A developer has built an open-source Telegram bot that uses Google's Gemma model running locally to read and understand bills, then sends persistent reminders through Temporal's durable workflow system until the bill is paid. Sentry is used for agent tracing without exposing bill data. The project is part of a weekend coding challenge and is being shared openly with the developer community.

  17. 17
    Apple Mac Studio with M5 Ultra runs frontier AI models locally▼Apple Mac Studio (M5 Ultra) Review: Unlimited Power The Mac Studio can run frontier-level AI language models locally. ItMmastodonTechnology22 d ago

    A new review of Apple's Mac Studio with the M5 Ultra chip says the desktop can run frontier-level AI language models locally, calling it a preview of what's to come. The Wired verdict, summarised as 'unlimited power', is drawing attention for suggesting high-end local hardware can now handle AI workloads previously reserved for cloud data centres.

  18. 18
    The Exercise Coach Brings AI-Assisted Workouts to Jacksonville▼AI-assisted fitness and workouts now available at The Exercise Coach in Jacksonville✉newsHealthFitness1 d ago

    The Exercise Coach, a fitness studio franchise, has introduced AI-assisted training at its Jacksonville location. The technology personalizes strength-training workouts, adjusting exercises to each client's ability and progress. The rollout highlights a broader trend of artificial intelligence entering the fitness industry, with local media noting the studio's machine-guided, time-efficient workout model now paired with AI-driven coaching tools.

  19. 19
    PewDiePie Says OpenAI Banned Him Twice While Building His Own AI●PewDiePie Says OpenAI Banned Him Twice While He Built Ajax, His Own Local AI Model✉newsTechnologyGadgets47 min ago

    YouTuber PewDiePie says OpenAI banned him twice while he was developing Ajax, his own local AI model. The claim, reported by Gadget Review, adds to ongoing attention around the creator's recent move into self-hosted technology, with audiences discussing both the bans and his decision to build an independent offline alternative rather than rely on mainstream AI services.

  20. 20
    Framework Desktop with 192GB memory opens pre-orders●192GB Framework Desktop open for pre-orderYhnLifeHome & Garden168 h ago

    Framework has opened pre-orders for its Desktop DIY configuration built around AMD's AI Max 400 platform, with configurations supporting up to 192GB of unified memory. The unusual memory capacity, rare in a compact desktop, is drawing attention from developers and enthusiasts running local AI workloads, who see it as a flexible alternative to traditional mini PCs.

  21. 21
    Citi: Open-Source AI Model Threat to Frontier Revenues Has Peaked▼[Major Bank] Citi: Impact of Open-Source Weight Models on Frontier Revenues Has Passed Local Peak; Capability Gap Widens Again✉newsTechnologySoftware1 d ago

    Citigroup analysts argue that the revenue pressure open-source weight models once placed on frontier AI developers has passed its local peak, and that the capability gap between leading closed models and open alternatives is widening again. The note, circulating on financial news feeds, suggests investors may reprice AI lab revenues as closed frontier models regain their technical lead.

  22. 22
    Ten-minute seated pose sketch shared by German drawing studio●Sitzende, Fineliner und Fasermaler auf Papier, Pose zehn Minuten. Modell: Jana # aktzeichnung # schnellestudien # art #MmastodonCultureArt412 h ago

    Atelier am Kirschgarten, a drawing studio in Germany, shared a quick life-drawing study of a seated figure by model Jana, drawn on paper with fineliners and felt pens in a ten-minute pose. The piece is part of the studio's regular figure-drawing sessions and quick-sketch practice, shared with the online art community alongside tags for traditional, human-made artwork.

  23. 23
    Apple overhauls macOS Full Disk Access to curb AI agents▼Apple is updating macOS security by overhauling its Full Disk Access permission model in direct response to AI agents. TMmastodonTechnology21 d ago

    Apple is revamping macOS's Full Disk Access permission model in response to the rise of AI agents. The new approach is designed to stop agentic apps from demanding broad, persistent access to sensitive data such as local files, Mail stores, iMessage databases and browser histories. Observers see it as a direct acknowledgement that AI software is reshaping what desktop security rules must protect against.

  24. 24
    Redis creator launches ds4 for running LLMs locally●From the creator of Redis; run LLM locally with ds4 Article URL: https:// dwarfstar.sh/ Comments URL: https:// news.ycomMmastodonTechnology220 h ago

    A new tool called ds4, promoted as coming from the creator of Redis, lets users run large language models on their own machines. The project is being shared on developer forums, where early readers are weighing its promise of private, local AI inference. Details on features and licensing remain thin, and discussion is just beginning.

  25. 25
    Developer Breaks Down llama.cpp Configuration for Qwen 3.8B●Understanding My llama.cpp Qwen 3.8 Configuration I've been tuning llama.cpp for local AI development, and the command lMmastodonTechnologyAI21 d ago

    A developer has published a parameter-by-parameter walkthrough of their llama.cpp setup for running the Qwen 3 8B model locally, explaining what each command-line flag does and how the options are tuned for maximum performance on their hardware. The guide is aimed at people running AI models on their own machines, where cryptic command-line options often make local inference setups hard to understand and reproduce.

  26. 26
    UC Santa Cruz's Adam Smith on local small language models●Adam Smith from UC Santa Cruz joins us to discuss local Small Language Models (SLMs) and building open, autonomous toolsMmastodonTechnologyAI21 d ago

    Adam Smith of UC Santa Cruz is discussing the case for running small language models locally rather than relying on large cloud providers. He presents BayLeaf AI, described as a counterplatform, along with the concept of "transagency" — a human-agent collaboration model he likens to the relationship between a driver and a car. The conversation also covers context distillation and practical approaches to building open, autonomous AI tools that users control themselves.

  27. 27

    The GLM 5.3 Flash model is reportedly capable of running at frontier-level performance on a pair of Nvidia DGX Spark desktop systems, according to the claim drawing attention online. The setup suggests advanced AI inference can now be achieved on compact, relatively affordable local hardware rather than large data centre clusters. Commenters are discussing the implications for accessible high-end AI.

  28. 28
    PewDiePie builds local AJAX AI model after OpenAI ban●Pewdiepie builds smaller, local AJAX AI model after OpenAI ban controversy✉newsTechnologyAI9 h ago

    YouTuber PewDiePie has built a smaller, locally run AI model called AJAX, following a controversy involving a ban from OpenAI. The model runs on local hardware rather than cloud services, reflecting a shift toward self-hosted AI tools. The story is being covered by tech outlets, with attention focused on both the OpenAI dispute and the move to independent, offline AI development.

  29. 29
    NVIDIA DGX Spark 64GB expands local AI development options●NVIDIA DGX Spark 64GB Gives Developers More Ways to Build and Scale Local AI✉newsTechnologySoftware1 d ago

    NVIDIA has announced the DGX Spark with 64GB of memory, a compact AI development system aimed at giving developers more ways to build and scale AI applications locally. The company says the machine lets developers prototype, fine-tune and run AI models on their desktop without relying on cloud infrastructure.

  30. 30
    PewDiePie Says OpenAI Banned Him Twice Over Local AI Model●PewDiePie Claims OpenAI Banned Him Twice Over Local AI Model𝕏xSE572 d ago

    YouTuber PewDiePie, real name Felix Kjellberg, claims OpenAI banned his account twice, which he attributes to his use of local AI models on his own hardware. The claim, made publicly by the creator himself, has drawn attention from tech communities debating platform moderation and the push toward self-hosted AI. OpenAI has not publicly commented on the alleged bans.

  31. 31
    Local AI decision model Bespoke Nimble draws experimenter interest●I’ve been experimenting with Bespoke Nimble, a local decision model running through Ollama. It takes evidence, a questioMmastodonTechnologyAI11 d ago

    A developer is testing Bespoke Nimble, a small decision model run locally through Ollama. The model takes evidence, a question and a set of allowed answers at request time, meaning the same model can handle many classification tasks without retraining. The author is comparing it with another model called Jev in a write-up, and interest centres on whether compact local models can replace task-specific trained classifiers.

  32. 32
    Anthropic proposes opt-out AI training rules for Australian content●TL;DR: AI company Anthropic calls for an opt-out model for Australian content to train its models, while ABC and SBS demMmastodonTechnology32 d ago

    Anthropic has told an Australian review that AI firms should be able to use locally published content for training unless creators opt out. The proposal puts the company at odds with Australian broadcasters ABC and SBS, who are demanding strict regulations to protect journalism and ensure media organisations are fairly compensated when their work trains AI models.

  33. 33
    TensorFold claims up to 3x faster LLM inference on Mac and DGX Spark●シタン先生もpythonについて話していました Mac・DGX SparkでLLM推論を最大3倍高速化する「TensorFold」の概要|npaka https:// note.com/npaka/n/n3d3e09549bdd # AppMmastodonWorld31 d ago

    A new tool called TensorFold is being described as able to speed up LLM inference by up to three times on Apple Macs and Nvidia's DGX Spark hardware. A Japanese-language explainer by npaka on Note is circulating, and comments reference discussions of Python in relation to the tool. The claim is drawing attention among AI developers interested in running large language models locally.

  34. 34
    Bilibili Open-Sources Translation Model Family Covering 150 Languages▼Bilibili Open-Sources Index-Translate, a Qwen3.5-Based Translation Model Family for 150 Languages✉newsTechnologySoftware2 d ago

    Bilibili has open-sourced Index-Translate, a family of translation models built on Alibaba's Qwen3.5 that supports 150 languages. The release puts a large multilingual translation capability into open weights, letting developers run and fine-tune it themselves. The move adds to a growing wave of Chinese tech firms releasing open-source AI models and could draw interest from localization and machine translation developers.

  35. 35
    Using a local LLM to clean up a full hard drive●I gave my local LLM a nearly-full SSD and told it to find everything I could safely delete✉newsTechnologyAI3 d ago

    A tech writer describes running a locally hosted large language model on a nearly full SSD, asking it to identify files that could be safely deleted. The piece highlights a practical, off-cloud use of local AI: letting the model scan the drive and suggest disk cleanup targets, reflecting growing interest in running LLMs directly on personal hardware for everyday tasks.

  36. 36
    Anthropic finds Zhipu's GLM-5.3 nearly matches Claude in cyber exploits●Anthropic evaluiert Zhipus Open-Weight-Modell GLM-5.3: Es generiert Cyber-Exploits nahe am Niveau von Claude Mythos. FürMmastodonTechnologyAI13 d ago

    Anthropic has evaluated Zhipu's open-weight model GLM-5.3 and found it generates cyber exploits close to the level of its own Claude Mythos model. At a reported cost of about 20.40 dollars per Chrome attack, local inference on security tasks already looks highly competitive, fueling debate over open-weight AI models reaching frontier capabilities in offensive cyber operations.

  37. 37
    Frustration with AI chief fuels local model push●Because fuck Dario Amadeus Mozart # ai # llm # local # localModels # yarrMmastodonTechnologyAI03 d ago

    A blunt post aimed at Dario Amodei, the head of AI company Anthropic, is circulating alongside hashtags about running large language models locally. The message pairs an insult with calls for local models and piracy ('yarr'), reflecting a sentiment among some users who oppose closed, corporate-controlled AI and prefer open or self-hosted alternatives they can run on their own machines.

  38. 38
    Local LLM helps hobbyist code a handy tool●Okay my local LLM helped me yesterday to vibecode something really handy. To be honest it did most of the heavy regex liMmastodonTechnologyAI23 d ago

    A developer says a locally run large language model helped him build a genuinely useful script, handling most of the tricky regex work in what he calls vibecoding. He now wants to release the tool publicly but is unsure how to do so on his Codeberg account without violating its terms, and is asking others for advice on the right way to share it.

  39. 39
    AMD driver update boosts Radeon AI performance up to 23%●🤖 AMD boosting AI/LLM performance for Radeon iGPUs as much as 18~23% with Linux 7.4 submitted by /u/Fcking_Chuck [link]MmastodonTechnologySoftware04 d ago

    AMD is delivering significant AI and large language model performance gains for its Radeon integrated graphics, with improvements of roughly 18 to 23 percent arriving via the Linux 7.4 driver. The gains matter for users running AI workloads on budget and portable systems that rely on integrated GPUs rather than discrete graphics cards. Linux users and AI enthusiasts are discussing what the update means for local LLM performance on AMD hardware.

  40. 40
    Running Xiaomi's MiMo Qwen 9B Distill on a 16GB MacBook Pro●This is the third video already in the series I started this month. I am trying to share my... # ai # programming # tutoMmastodonTechnologySoftware33 d ago

    A developer is demonstrating Xiaomi's MiMo V2.6 Qwen 9B distill model running locally on a 16GB MacBook Pro, the third installment in a tutorial series launched this month. The series covers AI, programming and agents, with an emphasis on inclusive, community-oriented learning. Interest centres on whether mid-range laptops without high-end GPUs can now run capable open-weight language models.

Repos