Skip to main content

Posts

IBM Moves Its Bob Coding Agent Inside the Firewall

Plenty of enterprises want AI coding agents. Fewer are willing to send their source code to someone else’s cloud to get them. That tension has slowed adoption at banks, insurers, government agencies and other organizations that build software under strict rules about where code and data can live. IBM is betting a self-hosted option will help close the gap. This week, the company announced that IBM Bob, its agentic software development platform, can now run on-premises, in private clouds, in sovereign clouds and in fully air-gapped environments. IBM made Bob generally available as a SaaS offering in April. At the time, it said on-premises deployment would come in a future release. That release is now here. Bob is meant to do more than complete code. IBM pitches it as a partner across the software development lifecycle, from planning and design through coding, testing, deployment, and modernization. It coordinates specialized agents for code, tests, documentation, and pipelines. ...
Recent posts

AWS AI Agent Surfaces Recommendations to Optimize Cloud Computing Environments

Amazon Web Services (AWS) today made available a preview of an artificial intelligence (AI) agent that surfaces recommendations to optimize cost, security, performance and resilience based on the business objectives an organization defines. Jill Fariss, vice president of AWS Support, said the AWS Well-Architected Agent generates code that can be implemented via a command line interface (CLI), an infrastructure-as-code tool or some type of runbook automation. The overall goal is not to eliminate the need for humans to manage IT infrastructure but rather make it simpler for engineers and IT administrators to manage workloads at much higher levels of scale, added Fariss. That capability is critical because most organizations today simply don’t have enough IT staff to effectively manage and optimize the workloads they currently have deployed, an issue that is only going to be further exacerbated as advances in AI make it possible to build and deploy even more applications, she noted. Th...

Introducing Futurum Media: Why Futurum, Why Now?

I am a builder. My career has spanned the rise of the commercial internet, web hosting, application service providers, managed infrastructure, security, DevOps, cloud native computing, platform engineering and now AI. Through all of it, I have been a builder and a leader. I didn’t think, after all these years, I would be building again quite like this. But in the age of AI, we can all be builders. I am happy to be doing this. Today, The Futurum Group is launching Futurum Media , bringing Techstrong, Tech Field Day and Visible Impact together into a go-to-market execution business. It becomes the umbrella for Futurum’s editorial properties and media services, backed by the broader organization’s research, intelligence and expertise. For those of you who read DevOps.com and our other publications, participate in our programs or work with us, here is why I believe this is a significant next step. A Community Worth Building On DevOps.com grew alongside a community changing how softw...

Survey Surfaces Sharp Increase in Amount of Code Written by AI

A global survey of 705 developers and IT leaders finds 42% of respondents reporting that artificial intelligence (AI) now writes at least half their code, with only 21% of developers now spending more than half their week writing new code from scratch. Conducted by BairesDev, a provider of software development services, the survey also finds nearly 80% of developers now spend less than half their week coding. As a result, developers on average are saving 13 hours a week on coding, allowing them to devote more time to reviewing AI output (67%) and debugging it (52%). Developers are also now spending, on average, nine hours a week learning AI tools and new technologies. That investment appears to also be paying off, with AI tool fluency (29%), system architecture (20%), and human skills (15%) expertise driving pay increases for developers, the survey finds. A full 86% of respondents also report they now find their role in their organization more fulfilling, according to the survey. I...

A Semicolon in a Branch Name Was All It Took to Steal an AI Agent’s GitHub Token

AI coding agents don’t just suggest code anymore. Tools like OpenAI’s Codex spin up a real container, clone a real repository, and authenticate with a real GitHub credential to get the job done — which means every agent your team wires up is also a new privileged identity, holding real access, running with comparatively little of the scrutiny a human with that same access would get. A Branch Name Was the Whole Attack In March, BeyondTrust’s Phantom Labs disclosed a critical command injection vulnerability in Codex. The flaw was almost absurdly simple: when Codex creates a task container, it passes the target branch name into a shell command without sanitizing it first. Characters like ; && | $() and backticks get interpreted literally by Bash. The proof-of-concept needed nothing more exotic than that. Set the branch to main , append a semicolon to terminate the intended git command, then inject a second command that writes the output of git remote get-url o...

Perforce Applies Machine Learning to Generate Synthetic Data for App Testing

Perforce Software has added a tool that leverages artificial intelligence (AI) to make it simpler for application development teams to generate synthetic data for application testing purposes. Mayank Ahluwalia, a senior product manager for Perforce Delphix, said Delphix Synthetic Data makes use of machine learning algorithms to generate synthetic data for specific use cases. It automatically identifies data structures, relationships, and business context across multiple sources. Historically, DevOps teams would have had to manually provide access to those data sources using some type of legacy tool, he added. The overall goal is to limit or eliminate the need to give application development teams access to production data in order to test an application, noted Ahluwalia. At the moment, providing those teams with access to data needed to run tests has become a bottleneck that can be eliminated by using a self-service platform for generating synthetic data, he added. That capability...

AWS Benchmark Aims to Reduce Number of False Positives Found by AI Vulnerability Scanners

Amazon Web Services (AWS) has developed a benchmark that can be used to test whether a model can distinguish real vulnerabilities from code that looks risky but is actually safe. The Deception Benchmark was created following an evaluation of the capabilities of 12 models from five different providers. In all, the benchmark includes 14,822 samples of code built using 16 different languages spanning more than 70 Common Weakness Enumeration (CWE) categories. Each sample is run through an adversarial loop to generate code, test it against frontier models, harden, repeat. If a model gets it right easily, the sample is removed. According to the benchmark, every AI model has the same fundamental issue. While they identify up to 95% of real vulnerabilities, they also flag 41 to 99% of safe code. Proof-of-exploit prompting can cut false positives by 17 to 74 percentage points but misses 7% to 44% of real vulnerabilities. The environment-gated challenges are worse: models flag the code and ig...