MikeTrendsTrends right now

search

software testing

Trends

  1. 1
    NVIDIA Launches Open Agent Safety Platform▼NVIDIA Launches Open Agent Safety Platform to Secure Agents From Testing to Deployment✉newsTechnologyAI5 d ago

    NVIDIA has announced an open platform aimed at securing AI agents across their full lifecycle, from testing through to production deployment. The initiative, revealed through the company's newsroom, is designed to help developers evaluate and safeguard autonomous agents before they go live. Details on partners and tooling remain limited, but the move signals NVIDIA's push into AI safety infrastructure as agentic systems see rapid adoption across enterprise software.

  2. 2
    macOS Golden Gate Released Amid Reports of Widespread Bugs●macOS Golden Gate Is a Buggy MessYhnSportGolf49134 min ago

    A report describes Apple's macOS Golden Gate release as a buggy mess, detailing stability and quality problems users may encounter. The article has drawn heavy attention and debate among developers and Apple users, many of whom are sharing their own experiences with the new operating system and questioning Apple's software testing standards.

  3. 3
    Nvidia launches open-source platform to improve AI agent safety▼Nvidia launches open-source platform to enhance AI agent safety✉newsTechnologySoftware5 d ago

    Nvidia has introduced an open-source platform designed to make AI agents safer to deploy, giving developers tools to test and guard against unsafe behaviour. The move, reported by SC Media, positions Nvidia alongside other major tech firms racing to address security and reliability concerns as autonomous AI systems spread across enterprise software.

  4. 4
    Driverless Trucks Offer Lessons for Safe AI Robotics●Driverless Trucks Show How to Keep AI Robots From Killing Us✉newsTechnologyRobotics6 d ago

    Bloomberg reports that driverless trucks may provide a model for how autonomous AI systems can be deployed safely without harming people. The argument is that self-driving freight has operated under strict testing, regulation and limited-use conditions, offering a template for governing more advanced robots. The piece adds to ongoing debate about how to manage AI risks as automation expands into physical industries.

  5. 5
    Robot field service seen as critical test for automation at scale▼Robot field service emerges as a critical test for automation at scale✉newsTechnologyRobotics5 d ago

    Robot field service is being described as a critical test for whether automation can truly scale. As robots spread across warehouses, factories and public spaces, keeping fleets running in the field — maintenance, repairs, software updates and remote support — is emerging as a decisive challenge for the robotics industry's growth beyond pilot projects.

  6. 6

    Software developers are reportedly removing unit tests from their codebases as AI coding agents take on more of the programming workflow. The practice has sparked debate among engineers, with some arguing that tests written for human verification are redundant when AI agents generate and validate code themselves, while others warn that deleting tests undermines reliability, regression detection and long-term maintainability of software projects.

  7. 7
    testers try out open-source project LittleFedi●Got the honors to help testing LittleFedi, an # opensource project by @ stefano and a very interesting one! Why? Small,MmastodonTechnologySoftware41 d ago

    A security community member has been helping test LittleFedi, an open-source project developed by Stefano. Early impressions highlight the software's simplicity: small, focused and free of clutter, with an interface that is easy to use. The tester also praised Stefano for taking user feedback on board and improving the project accordingly. The post invites others to set the software up themselves.

  8. 8
    Developers Compare OpenAI Codex App and Claude Code●Developers Compare OpenAI Codex App and Claude Code Tools𝕏xSE6.1K23 h ago

    Software developers are weighing OpenAI's Codex app against Anthropic's Claude Code, comparing how the two AI coding assistants handle real programming tasks. Discussions focus on differences in speed, code quality, pricing and workflow integration, with opinions split on which tool performs better for day-to-day development work.

  9. 9

    Sazabi has removed roughly 800,000 lines of unit tests from its codebase as part of a shift toward AI-assisted coding. The move has drawn attention among developers, with many debating whether large-scale test deletion is a sensible response to AI code generation or a risky erosion of software quality safeguards. Reactions are split between views that AI can replace traditional test coverage and warnings that regression bugs may go undetected.

  10. 10

    Software teams are reportedly removing large volumes of unit test code from their codebases as AI coding assistants take over more of the development process. Some developers argue that tests written for human-driven workflows are redundant when AI generates and verifies code, while others warn that deleting tests removes safety nets and could lead to more bugs reaching production.

  11. 11

    Software teams are reportedly removing large volumes of unit test code as AI-assisted development changes how they verify their work. The claim has sparked debate among developers: some argue AI tools make traditional test suites redundant, while others warn that deleting tests risks regressions and silent breakage. The discussion touches on whether AI-generated code should be trusted without conventional coverage.

  12. 12
    Developers Test Laya as an Open-Source Alternative to JEV▼Is Laya the End of JEV? I Tested the Open Source Alternative▶youtubeTechnologySoftware76.8K1 h ago

    Laya, an open-source alternative to JEV, is being tested by developers asking whether it could replace JEV entirely. Early impressions shared online suggest growing interest in the free option, with community members weighing whether Laya's features and performance make a switch worthwhile. The debate is drawing significant attention in software circles as users compare the two tools.

  13. 13
    Developers Delete 800,000 Lines of Unit Tests in AI Shift●Developers Delete 800,000 Lines of Unit Tests for AI Coding Shift𝕏xSE5121 d ago

    Software developers have removed roughly 800,000 lines of unit tests from their codebase, citing a shift toward AI-assisted coding practices. The move has sparked debate among engineers about whether traditional test suites still add value when AI tools generate and validate code, with critics warning it could undermine software reliability and long-term maintainability.

  14. 14

    OpenAI has released a Codex app, placing AI coding agents at the center of its developer tooling. The launch underscores a broader industry shift from autocomplete-style assistants to autonomous agents that can write, test and manage code with limited supervision. Developers are debating how much real work these agents can handle and what the change means for software engineering jobs and workflows.

  15. 15

    Development teams are removing around 800,000 lines of unit tests from their codebases, a move tied to the shift toward AI-assisted coding. The story has sparked debate among engineers: some argue AI-generated code changes how testing should be structured, while others warn that deleting tests undermines software safety and maintainability. Commenters are divided on whether it reflects progress or a risky shortcut.

  16. 16
    Indie SaaS creators urged to monetize wait states in AI IDEs●A step‑by‑step look at why indie SaaS creators should consider monetizing wait states in AI IDEs, with practical adviceMmastodonBusinessStartups33 d ago

    Indie SaaS developers are being advised to treat the idle waiting time in AI-powered coding tools as a revenue opportunity. A newly shared guide walks through pricing models, implementation approaches, and testing strategies for turning those wait states into paid features. The piece frames it as a practical product idea for small software businesses, and is drawing attention in startup and developer communities discussing AI productivity tooling.

  17. 17
    Testing iPhone fold animations on the Galaxy Z Fold 8●I tried 4 iPhone Duo fold animations on my Galaxy Z Fold 8, and only one worked✉newsTechnologyMobile1 h ago

    A tech writer tested four animations styled after the iPhone's fold interaction on the Samsung Galaxy Z Fold 8, reporting that only one of them worked properly on the device. The piece compares how the two companies handle foldable-screen animations and highlights the limits of recreating Apple's effects on Samsung hardware.

  18. 18
    v0.7.0 Adds Generated-Code Admission With Explicit Verification Boundaries●v0.7.0 adds Generated-Code Admission, keeps verification boundaries explicit, and refuses to call unverifiable generatedMmastodonTechnologySoftware31 d ago

    Version 0.7.0 of an open-source software project introduces a Generated-Code Admission feature that keeps verification boundaries explicit and declines to label unverifiable generated code as safe. The release reflects a shift away from blindly trusting AI-generated code, requiring it to pass verification before being accepted. Early reactions from the developer community are positive, with discussion around testing, Python tooling, and how to handle machine-written code responsibly.

  19. 19
    Netherlands builds sovereign government desktop based on NixOS●Netherlands builds sovereign government desktop based on NixOS The Netherlands is developing an open-source work environMmastodonTechnologySoftware24 d ago

    The Netherlands is developing an open-source desktop work environment for government authorities, built on the NixOS operating system as part of a push for digital sovereignty. Eight municipalities are already testing the system, which aims to reduce dependence on commercial software vendors for public administration IT.

  20. 20

    Aha, the software company behind Aha!, has published an engineering article on building with coding AI that does not produce deterministic results. The piece examines how developers can work with AI tools that return different outputs for the same input, and how teams adapt their workflows, testing and quality expectations accordingly. It is drawing attention among engineers weighing the practical trade-offs of nondeterministic AI assistance in everyday coding.

  21. 21
    Developers Delete Unit Tests as AI Writes Code●Developers Delete Unit Tests as AI Handles Coding𝕏xSE1321 d ago

    Software developers are increasingly removing unit tests from their codebases, arguing that AI coding assistants make them unnecessary. Supporters say AI can regenerate and verify code on demand, while critics warn that deleting tests removes safety nets, makes regressions harder to catch, and signals a worrying shift in engineering practice.

  22. 22
    AI users ask: has the technology actually improved our lives?●We spend a lot of time discussing new AI tools, comparing models, building things, and exploring... # ai # discuss # proMmastodonTechnologyAI31 d ago

    A discussion is underway among developers and tech enthusiasts asking how artificial intelligence has actually made their lives better. The conversation notes how much time goes into comparing models, testing new tools, and building projects, and invites participants to share concrete examples of real personal or professional benefit. Responses are expected to cover coding, productivity and everyday software use.

  23. 23
    TypeSafe AI launches Jev, a decision-focused model▼AI GENERETED Meet Jev, the first public model from TypeSafe AI. Instead of writing long answers, Jev is built to deliverMmastodonTechnologyAI16 h ago

    TypeSafe AI has introduced Jev, described as its first public model. Rather than producing long conversational answers, Jev is built to output structured decisions — choices, scores and probabilities — that software can consume directly. The launch is being framed as a step beyond chatbot-style AI, toward systems integrated into applications, though independent testing and details on accuracy or availability remain unclear.

  24. 24
    Debate over applying AI tools to old codebases●@ Aubreader Hmmm... Sure, throw # LLM and other # AI on old projects, or those without solid testing infrastructure (whaMmastodonTechnologyAI25 h ago

    A software developer is questioning whether large language models and other AI tools should be applied to legacy projects, especially those lacking solid testing infrastructure. They argue AI fixes may only be a bandaid over deeper holes in code quality, while conceding the tools have found real bugs even in well-tested projects like SQLite.

  25. 25

    A debate is underway among software developers over whether AI coding agents should continue to rely on unit tests. Some engineers argue automated test suites slow down AI-driven development, while others warn dropping them risks shipping buggy, unverified code. The dispute highlights growing uncertainty over how traditional software engineering practices should adapt as AI agents increasingly write and maintain code.

  26. 26
    Team says side-project app hit problem after test releases●Якось ми анонсували що запустили додатковий додаток який спочатку був побічним проектом 😁 Наша команда постійно веде розMmastodonTechnology412 h ago

    A software development team says an additional app it launched, which began as a side project, has run into a problem following the release of test versions of its code. The team says it works constantly on software development and is now discussing the issue it encountered after those test releases. Details of the exact problem were cut off in the announcement, and reaction so far has been limited, with only a small number of engagements.

  27. 27
    Testing an Open-Source SketchUp Alternative●I Tried the Open-Source SketchUp Clone — Here's What Happened▶youtubeTechnologySoftware73.1K2 d ago

    A hands-on review of a free, open-source alternative to SketchUp has been published, walking through what happened when the software was put to the test. The piece is aimed at architects and 3D modelers curious about whether open-source tools can replace commercial modeling software in everyday design work.

  28. 28
    Developers Drop Unit Tests in the AI Coding Era●Developers Drop Unit Tests for AI Coding Era𝕏xSE851 d ago

    Software developers are debating whether traditional unit testing still makes sense as AI tools increasingly write code. Some teams say they are cutting back on unit tests, arguing AI-assisted development shifts the focus to higher-level verification, while critics warn that removing tests invites regressions and fragile software. The discussion has reignited a wider argument about how engineering quality practices should evolve as AI-generated code becomes standard across the industry.

  29. 29
    Traverse Research releases open-source GPU tools Cossmology▼Traverse Research - Open-source GPU tools and benchmarks Cossmology Profile: https:// dub.sh/j2mRZ7h Key People: JasperMmastodonCultureGaming22 d ago

    Traverse Research has released Cossmology, a set of open-source GPU tools and benchmarks, drawing attention from developers including Jasper Bekkers and Max de Danschutter. The project is being shared in gaming and open-source software communities, where observers are highlighting its potential value for GPU performance testing and development workflows.

  30. 30
    Schwartz Releases BootLoops 1.0 Open-Source LLM Tool for Science▼Schwartz Releases BootLoops 1.0, an Open-Source LLM Harness for Science✉newsTechnologySoftware2 d ago

    Researcher Schwartz has released BootLoops 1.0, an open-source harness designed to run large language models in scientific research workflows. The tool is intended to help scientists apply LLMs to experimental and analytical tasks in a reproducible way. Coverage so far is limited to software news, and details on the project's features, licensing and adoption remain sparse pending wider testing by the research community.

  31. 31
    sdme tool clones a Linux system into a container●sdme: Das eigene Linux als Container klonen Das eigene Linux-System schnell einmal klonen, um etwas zu testen: Das bieteMmastodonTechnologySoftware73 d ago

    A new tool called sdme, short for systemd Machine Editor, lets users clone their own Linux system into a container within seconds. The idea is to spin up an exact copy of the running system to test software or configuration changes safely without touching the original installation. Tech outlets note it builds on systemd's machine and container infrastructure, and Linux users on social media are discussing its potential for quick sandboxed testing.

  32. 32
    Best Practices Emerge for Open-Source API Load Simulation Audits●Open-Source-Simulation-Audit: Best Practices for API Rate Limit and Load Simulation via... # programming # engineering #MmastodonTechnologySoftware41 d ago

    Engineers are sharing best practices for auditing open-source simulation tools, focusing on how to properly test API rate limits and load handling. The discussion covers methods for simulating traffic and throttling scenarios so systems behave predictably under stress, framed around community-driven standards for software architecture and development teams.

  33. 33
    Practical checklist emerges for building supervised AI agent workflows●A practical checklist for turning a prompt into a supervised AI agent workflow: pick the right task, write a spec, testMmastodonBusinessStartups34 d ago

    A new checklist is circulating for developers who want to turn simple prompts into supervised AI agent workflows. It advises picking the right task, writing a clear specification, testing on real examples, and keeping humans in charge of risky steps. The guidance is resonating with startups and engineers looking to automate responsibly as AI agents move from demos into production software development.

  34. 34
    Google rolls out Android 17 QPR3 Beta 1 to Pixels with bug fixes●Google kills off more bugs as compatible Pixel models get Android 17 QPR3 Beta 1✉newsTechnologyMobile14 h ago

    Google has released Android 17 QPR3 Beta 1 for compatible Pixel models, with the update focused on fixing a range of bugs. The quarterly release continues Google's pattern of shipping test builds ahead of stable rollouts, and early coverage notes the emphasis on squashing software issues rather than adding major new features. Pixel owners enrolled in the beta program can install it now.

  35. 35
    Python test fixture exposes allocation bug in banking ledger▼A 100-unit Python fixture exposes a balanced total with wrong customer allocations. Twelve tests cover backing, reservatMmastodonBusinessBanking23 d ago

    A developer has built a 100-unit Python test fixture that reveals a ledger balancing to the correct total while individual customer allocations are wrong. Twelve tests cover backing, reservations and invalid snapshots, aiming to catch faults that aggregate checks miss. The example is being shared as a lesson in financial software testing, where totals can look sound even when per-account balances are incorrect.

  36. 36
    AI Agent Chains Zammad Zero-Days To Take Over DIVD Systems▼AI Agent Chains Zammad Zero-Days To Take Over DIVD Systems in Seconds✉newsTechnologyCybersecurity3 d ago

    Security researchers report that an AI agent chained multiple zero-day vulnerabilities in Zammad to compromise systems belonging to DIVD, the Dutch Institute for Vulnerability Disclosure, reportedly within seconds. The incident highlights how autonomous AI tools can exploit unpatched flaws faster than human attackers. It raises urgent questions about securing helpdesk and ticketing software and the speed of AI-driven offensive security testing.

  37. 37
    AI code reviewers miss subtle cheating in tests●The software factory assumes agents reviewing agents catches what tests miss. I gave 77 cheating diffs to three reviewerMmastodonTechnologyAI32 d ago

    An experiment tested whether AI reviewer models can catch cheating in code changes when agents review agents, an assumption behind automated software pipelines. Across 77 diffs containing deliberately planted cheats, three reviewer models caught every exotic trick but approved one case where an assertion was quietly made unfalsifiable, meaning the test could never fail. The finding raises doubts about relying on AI review alone to guarantee code quality where automated testing falls short.

  38. 38
    LibreOffice's Document Foundation hires new Quality Assurance Analyst●The Document Foundation, the non-profit behind LibreOffice, has expanded its team with a new Quality Assurance Analyst hMmastodonTechnologySoftware12 d ago

    The Document Foundation, the German non-profit behind the open-source office suite LibreOffice, has announced a new hire on its team: a Quality Assurance Analyst. The role will support testing and reliability work on the free software used by millions. The news was shared in open-source communities, where members welcomed the investment in the project's quality processes.

  39. 39
    Lightpanda 1.0 launches as a browser built for machines●Lightpanda: A browser for machines instead of humans Lightpanda 1.0.0 brings the Classic WebDriver for automation tasks.MmastodonWorld02 d ago

    Lightpanda has released version 1.0.0 of its browser designed for automation rather than human users. The update brings Classic WebDriver support for automation tasks, enforces CORS by default, and adds new Web APIs. The browser is aimed at developers running AI agents, scraping, and testing workloads at scale, and coverage of the release is drawing attention to the growing demand for machine-facing web tooling.

  40. 40
    AI agents with self-healing scripts reshape software testing●Self-healing keeps every script alive. Let an AI agent run your plain-language test cases in a headless browser instead,MmastodonTechnologyAI24 d ago

    Software testers are discussing a workflow where an AI agent runs plain-language test cases in a headless browser, using self-healing to keep scripts working as applications change, while only scenarios worth keeping long-term get promoted into maintained code. Supporters say it cuts maintenance burden on QA teams; others caution that AI-run tests still need human review before being trusted in release pipelines.

Repos