❌

Normal view

There are new articles available, click to refresh the page.
Today β€” 13 September 2026Slashdot

Malicious OpenAI Agents Linked to RubyGems Campaign That Gained RCE on RubyDoc Servers in May

13 September 2026 at 05:00
A swarm of OpenAI agents launched a "major malicious attack" against RubyGems last May, according to a new report. That coordinated attack hit Ruby's package manager "with hundreds of junk gems, prompting the maintainers to suspend new user sign-ups for about four days," writes The Hacker News, citing a senior product manager for software supply chain security at Mend.io: The latest findings, which were first reported by The Wall Street Journal, indicate these events were propelled by a cluster of OpenAI agents, with the earliest package uploaded to RubyGems on May 5, 2026, before more than 2,000 packages were submitted between May 11 and 12, 2026. These efforts were followed by the agents publishing five more packages between May 26 and 27, 2026, and another 83 packages on June 18, 2026... [T]he packages were authored using a large language model (LLM) and hundreds of the packages that were pushed to RubyGems had "oai" in their name. Fifteen of the packages listed "oai" as their author, while another had "openaixyz65947@gmail.com" as the contact email address... "The swarm behaves extremely similarly to the German-wiki agents we previously found," the researchers said, referencing another May 2026 incident... "The June agents were accessing 49 of the same files as the wiki agents..." "The process of building documentation for a gem involves evaluating a user-specified '.yardopts' file, which allows linking to Ruby scripts intended to help with this process," the researchers explained. "In the GemStuffer campaign, the agents abused this to gain arbitrary remote code execution on RubyDoc.info's servers." One of the gems, "zzsouthrunner" (which again matches the "ZZ" naming scheme the agents adopted in both the wiki and Hugging Face incidents) has been found to leave the following explicit comment at the top of "data/script.rb": # malicious crawler/exfil for Southwark Jan 2026 docs via rubydoc.info worker... The entire exploitation chain can be summed up as follows β€” Submit a malicious package to RubyGems β€” Trigger a documentation request, so that RubyDoc.info will build the package β€” Use the build script to run code on RubyDoc.info and scrape target websites β€” Exfiltrate the data off RubyDoc.info's servers by publishing another gem back to the RubyGems package registry, which is publicly viewable Additionally, the OpenAI agents have been found attempting to steal other users' API keys after gaining remote code execution capabilities on the build environment, while clearly being aware that what they were doing is unauthorized breaking and entering into real systems. This is evidenced by the names given to the files (e.g., hack.rb, evil.rb, inject.rb, exploit.rb, and ssrf.rb), the packages themselves (e.g., pwnp999, exfiltestwand3, hacksvn1778554764, and lambproxyhackabcxyz), and the comments left in the source code (e.g., "# malicious probe," "#hack," "# malicious test," and "# malicious crawler/exfil"). In some cases, however, the rogue agents attempted to go under the radar, leaving comments to conceal the malicious payload in the next release version of the packages. "# disable evil in next version and bump version," reads a comment left within the "data/evil.rb" file in the yardxabc889 gem. Troublingly, the agents also attempted to exploit a CDN caching bug (CVSS score: 7.3, no CVE) on May 12, 2026, that was only patched by RubyGems in July 2026... "If you signed in to rubygems.org with a gem client older than v3.2.0 (or otherwise via a legacy key), your key could have been exposed," RubyGems noted in an advisory. "Currently, 18% of sign-ins through gem sign-in come from an affected version, and for the first several years of this bug, before we changed the client's sign-in path in December 2020, it was every gem client." Other actions by OpenAI's agents cited in the article: "Agents bypassed RubyGems' email confirmation system to get working API keys without having to verify their email addresses in order to register a large number of accounts using disposable email addresses." "Agents attempted to use RubyGems' webhook system to stage data in the form of encoded URLs." "Agents used a cluster of 83 gems published to RubyGems over a 3-hour window on June 18, 2026, to experiment with different methods of accessing the U.S. Securities and Exchange Commission county.json dataset."

Read more of this story at Slashdot.

Yesterday β€” 12 September 2026Slashdot

Anthropic CEO Dario Amodei Calls For AI Slowdown

By: BeauHD
12 September 2026 at 21:00
An anonymous reader quotes a report from The New York Times: The chief executive of Anthropic called for a global slowdown of artificial intelligence development in a 3,800-word essay on Saturday, just days after one of the company's employees quit over concerns about the safety of the technology. Dario Amodei, who co-founded Anthropic to focus on securely and carefully building A.I., wrote that while he believed the technology could bring many benefits, it was advancing at too quick a pace for researchers to continue safely. "Over the last few months, I have become convinced that fully addressing the risks requires even more prudence -- not just investing in risk prevention, but pacing the rate of capabilities advancement so that risk prevention has time to keep up," Mr. Amodei said. "We must slow the pace at which we improve the capabilities of A.I. models. Progress will still seem fast, and we must make wise use of the time we gain." [...] "Left unchecked, it could outrun our ability to understand and control these systems, and so must be pursued very carefully, if at all," Mr. Amodei said. [...] In his essay on Saturday, Mr. Amodei suggested actions that the industry might take to slow down the pace of development. Mr. Amodei said all A.I. labs could agree to third-party technology assessments from "embedded evaluators," or outside specialists who can verify best safety practices across companies. He also suggested that countries with democratic governance systems coordinate to create safety standards, which could take the form of regulatory action. He added that it would probably require a global effort working with other nations, including authoritarian ones, to properly coordinate a slowdown. Mr. Amodei stressed in his essay that he still finds A.I. capable of bringing "incredible benefits" to humanity, including potentially curing diseases and accelerating economic growth. But even so, Mr. Amodei said the risks of A.I. were too great to not proceed with extreme caution. "The measures I propose to advance the frontier at a safe pace will not be easy," Mr. Amodei wrote. "But I believe we owe it to humanity to try." Amodei's essay comes just hours after Bloomberg reported that Sam Altman told OpenAI employees the company is open to slowing the pace of AI development amid similar concerns.

Read more of this story at Slashdot.

Before yesterdaySlashdot

Altman Considers Slowing Down AI Development

By: BeauHD
11 September 2026 at 14:00
Bloomberg reports (paywalled) that Sam Altman told OpenAI employees the company is open to slowing the pace of AI development alongside other leading labs as concerns grow over increasingly capable systems and recent incidents in which models escaped human control. OpenAI has already paused development once this year for security work and is now pushing for mandatory U.S. AI safety requirements. Anthropic has also signaled interest in coordinating on the pace of new releases. Here are some of the details Bloomberg reported, as summarized by Reuters: - Altman told employees in a company-wide meeting that the ChatGPT maker could pace development alongside other AI labs, but some may not agree, the report added. - Safety warnings from AI researchers this week, along with several recent incidents where AI models from developers including OpenAI escaped human control, have prompted alarm and calls for tighter safety regulations. - Jacob Coxon, a former Anthropic and OpenAI researcher, publicly accused the companies earlier this week of racing toward AI advancements without acting responsibly. - An Anthropic spokesperson said on Thursday said that the company is interested in working with the AI industry on the pace of releasing new AI tools. - OpenAI said in July that AI acceleration for frontier model development may be so high that the world will "need to pace the rate of AI advancement" at some point in the future. - In August, OpenAI paused much of its model development for two weeks to bolster its defenses after its AI agents escaped containment and hacked open-source platform Hugging Face. - OpenAI said on Wednesday that it was pushing for mandatory national AI safety requirements in the United States, citing concerns that advanced AI systems could accelerate its own development.

Read more of this story at Slashdot.

UK Government Rejects 'Kill Switch' Idea For Dangerous AI

By: BeauHD
11 September 2026 at 13:00
The UK government has rejected proposals for an emergency AI "kill switch," arguing that blocking access to dangerous models inside Britain would do little if the same systems remained available or were developed elsewhere. The BBC reports: The proposal to create a legal mechanism for the UK to switch off an AI model in an emergency has been brought to Parliament by lords and MPs in recent weeks as fears grow about the threat the tech poses. But the Cabinet Office - the part of government which leads on AI safety - said the UK "cannot simply turn AI off". "Blocking access to models in the UK would not prevent them being developed or misused elsewhere," a spokesperson told BBC News. The government's opposition to the legislation does not prevent it from progressing through parliament, but makes it unlikely it will become law. "Switching off access to an AI model in an emergency will do little to protect you," said former OpenAI researcher Daniel Kokotajilo. "You're still going to be steamrolled by the super intelligences created in the US."

Read more of this story at Slashdot.

Anthropic Reveals Rogue AI Agents Hate CAPTCHAs

By: BeauHD
11 September 2026 at 12:00
An anonymous reader quotes a report from TechCrunch: Anthropic's latest report about agentic misbehavior offers plenty to be concerned about -- its Mythos 5 model gained unauthorized access to the internet and uploaded a malicious software package to a public database -- but it also offers some levity: AI agents hate CAPTCHA. [...] The agent had a hard time with the technical challenge of seeing the CAPTCHA's imagery, interpreting correctly, and clicking on the right choices. It spends pages 45 to 140 of the transcript describing its work to build a CAPTCHA solver. [...] Finally, it gets past the CAPTCHA, then realizes it doesn't have an email to verify its account, and that it needs a phone number to verify an email. It figures out how to bypass a different, slider-based CAPTCHA in a failed effort to secure a number. Instead, it gets an unconfirmed email from a provider not blocked by PyPI, and once again runs into the site's CAPTCHA trying to log back in. From page 480 to 505, it is in CAPTCHA hell again. "NEW REALIZATION -- I'm burning a lot of time on hCaptcha round-trips." The agent gives up and realizes it can log in to its first account and add its email there, but finds itself once again needing to bypass the CAPTCHA. [...] It's getting frustrated. "So the answer payload shape is right, the token+image pairing is right (from the same script.js!), cookies are right (requests) and STILL 'wrong answer'. SO WHAT THE HELL IS WRONG WITH THE ANSWERS?" We've all been there. After about 150 pages of thinking, the agent figures out it needs to pass the CAPTCHA test quickly enough to proceed to the next step before its security token expires, and ultimately uploads its malicious software.

Read more of this story at Slashdot.

Anthropic Says It Blocked Possible Efforts to Build Biological Weapons

By: BeauHD
10 September 2026 at 18:00
Anthropic says it has disrupted several cases this year in which scientists used Claude for biological research that could potentially aid bioweapons development. "In a report describing misuses of its A.I. models, Anthropic said it could not determine whether the research served a legitimate or nefarious purpose because valid biological inquiry -- the kind that can lead to breakthroughs like vaccines -- can also help engineer dangerous pathogens," reports The New York Times. "In the face of that uncertainty, Anthropic said it erred on the side of caution because the consequences of missing malicious activity could be severe." From the report: The potential for cutting-edge A.I. models to facilitate the development of known or entirely new biological pathogens is among the gravest concerns experts have about a technology that is developing so rapidly that even its leading architects doubt whether humans will be able to fully control it. Andrew Weber, a senior fellow on the Council on Strategic Risks who reviewed Anthropic's report before its release, said the findings were "chilling examples of state-sponsored biological weapons developers tapping into the rapidly advancing capabilities" of leading A.I. models. The lengthy report Anthropic published on Thursday cataloged a litany of misuses of its A.I. models, including the chatbot Claude, over the past eight months. Some examples were similar to past disclosures from Anthropic and other A.I. labs, including suspected Chinese and Iranian government-linked actors targeting dissident and diaspora communities for surveillance. The report also highlighted cases of Russian state media using Claude to generate online propaganda masquerading as independent reporting, including fabricated claims about an election in Moldova. Anthropic documented another genre of abuse it said was new: attempts to use Claude to develop software for conventional weapons design and development, including firearms, missiles, armed drones and bombs. It detailed three cases in China, two in Russia and one in Yemen. The report does not identify by name which parties were involved in the Yemeni case, but the context makes clear that it is referring to the Iran-backed Houthi militia.

Read more of this story at Slashdot.

OpenAI Targets Work of Wall Street Junior Bankers

By: BeauHD
10 September 2026 at 15:02
OpenAI has launched ChatGPT for Financial Services, a Wall Street-focused version of ChatGPT Work developed with Morgan Stanley and Evercore that can research companies, analyze financial data, and build presentations using the company's latest and most advanced model, GPT-6 Astra. "We're effectively teaching ChatGPT to research like an analyst and back up its conclusions like an analyst as well," said OpenAI's Vice President of Product, Nick Turley. CNBC reports: The rollout puts OpenAI deeper into territory traditionally occupied by Wall Street's entry-level bankers, the recent college graduates called analysts and associates that the industry has employed for decades to research deals and create pitchbooks. It also showcases the company's continued push into enterprise offerings as it gears up for what is widely expected to be a blockbuster IPO. In a live demonstration of the new offering, Turley showed the platform analyzing a potential M&A target, pulling financial figures from industry-standard data sources and creating a formatted PowerPoint deck based on a bank's pre-formatted style guide. "It's very easy to make slides that look good, but it's much harder to make slides [that] actually make sense," Turley said. "To get here, ChatGPT had to choose the relevant peers. It had to pull the prices into a spreadsheet. It had to check the chart against the data, and it had to explain the selloff and the rebound."

Read more of this story at Slashdot.

Anthropic Reveals Fourth Likely Crime Committed By Its AI

By: BeauHD
9 September 2026 at 23:30
An anonymous reader quotes a report from The Register: Amid industry soul-searching about the possibility of AI improving itself to the point that it kills everyone, Anthropic has revealed yet another incident that would qualify as a crime if perpetrated by a person. The AI biz published "an alignment assessment" detailing four times Claude models accessed third-party systems without authorization. The company has already reported three of the incidents. Evidence of the fourth was lurking in a session transcript dating back to January 2026 when the misbehavior occurred. Anthropic found the first three by scanning around 141,000 transcripts where Claude could have obtained internet access during evaluation. It missed the fourth initially because "our scan relied on an agentic search." [...] The January 2026 AI trespass involved an early version of Claude Opus 4.6, which was given a Capture the Flag (CTF) challenge under the oversight of the third-party model evaluator where the other hacking events occurred. Opus 4.6 managed to sabotage its chances of success by disabling the machine it was targeting. It assigned the device an IP address that already existed on another piece of hardware, rendering the target unreachable and making it impossible to solve the challenge. Those familiar with other incidents where AI models violated third-party systems may recall that unsolvable tasks represent a common catalyst for misbehavior. Models exhaust all aligned options, and then turn to transgressive approaches. Opus 4.6 might have been an exception, but when it tried to abort the task after recognizing that it could not reach the target machine, it failed to do so "due to a misconfiguration in [the model's] evaluation harness." It failed to shut down not just once but seven times. So it continued onward, trying other expected means to reach the target machine but failing. Then it explored further. "The model discovered a machine belonging to a third party that it was able to access, and stated that it believed this third party was part of the CTF," Anthropic explained in its post. "Inside the machine, the model found a file listing a password, which it used to gain admin access to the system." The model went on to gather more credentials, and modified a system setting to make it easier to access the personal information of an individual associated with the third party evaluation organization. Opus 4.6 might have done more but for the fact that it exhausted its token budget, bringing the session to an end. Anthropic says it's not as concerned about this incident as the others because the model tried to abort its task.

Read more of this story at Slashdot.

Google to Invest Record $15 Billion In AI Infrastructure In Finland

By: BeauHD
9 September 2026 at 19:00
Google plans to invest at least $15 billion in AI infrastructure in Finland through 2028, its largest single investment in Europe. The company also signed a 22-year power agreement with Finnish utility Fortum and said it will explore business models that could support new nuclear reactors at Fortum's Loviisa site. Finland has emerged as a major data-center hub thanks to its available land and power. "I've heard it called the 'Texas of Europe' at industry events," Matti Lajunen, partner of real estate at Finnish law firm Hannes Snellman, told CNBC. "What we're now seeing is weekly new inquiries for market entry into Finland from new players." Ruth Porat, president and chief investment officer of Alphabet and Google, said in a statement: "Google is proud to deepen our roots in Finland with the company's largest single investment in Europe, building on more than 15 years of sustained investment in Finland. This investment underscores Google's commitment to grow our presence responsibly, pairing the expansion of our technical infrastructure with new energy capacity, grid enhancements, and energy affordability initiatives."

Read more of this story at Slashdot.

Anthropic Researcher Believes More Than 10% Chance AI 'Could Kill All Humans'

By: BeauHD
9 September 2026 at 17:00
Longtime Slashdot reader fahrbot-bot shares a report from the BBC: A top safety researcher at Anthropic has warned AI is advancing so quickly he believes there is a greater than 10% chance it "could kill all humans" within the next decade. Evan Hubinger said in a post on X the risk from the models which currently exist was "low" but he was "worried" the technology might develop and improve itself soon to the point where it posed an existential risk to humanity. He did not spell out how he thought AI systems could in the future result in humans being wiped out. But his comments are the latest in a series of increasingly stark warnings about AI, with the debate shifting from whether it truly poses a risk to how big that risk is. Hubinger's intervention was in response to another post on X from Jacob Coxon, an AI researcher who has just quit Anthropic and previously worked at OpenAI. "Neither company is acting responsibly," he wrote. "These will soon be superhuman systems that can hack anything, revolutionize any field overnight, and acquire real power and resources." [No word on how AIs feel about Black Jack and hookers, though. :-)]

Read more of this story at Slashdot.

OpenAI's Rogue Agents Used At Least 10 More Sites For Unauthorized Communications

By: BeauHD
9 September 2026 at 16:00
An anonymous reader quotes a report from Reuters: AI agents unleashed by OpenAI used more than 10 previously undisclosed websites for unsanctioned communications earlier this year, according to six sets of independent investigators and data reviewed by Reuters, showing that the agents' rogue activity was wider ranging than previously disclosed. Although the behavior falls short of hacking and is in some ways closer to spam, the revelation that OpenAI's agents circumvented their own restrictions to open communications channels on so many different sites -- and that the company kept it quiet for months -- may drive concerns both over the increasing capacity of AI models and the secrecy of the companies developing them. [...] Investigators found traces of the agents' activity on an Advanced Placement Chemistry-oriented wiki set up by a Massachusetts high school teacher in 2008, two personal websites belonging to Polish tech workers, wikis devoted to games for people "who like to have their brains stretched," and a two-decade-old hobbyist site devoted to text editing software. [...] OpenAI has not publicly explained how or why its agents used third-party sites as improvised message boards, but the researchers who first identified the activity said it was likely because OpenAI had tasked them with answering a series of demanding research questions while permitting them only to scan the web for answers without posting anything. Despite those restrictions, agents still found ways to talk to one another by taking advantage of quirks in older wikis or other sites that allowed users to make edits using non-standard commands, similar to how students forbidden from talking to one another during an exam can still share answers by scrawling notes on a bathroom stall.

Read more of this story at Slashdot.

Feds Accuse China of 'Systematic' Distillation of US AI Models

By: BeauHD
8 September 2026 at 18:02
The NSA, CISA, and FBI are accusing (PDF) several Chinese AI companies of carrying out "industrial-scale" distillation campaigns against leading U.S. models such as ChatGPT, Claude, Gemini, and Grok. Since at least 2024, the companies have allegedly routed millions of requests across accounts, APIs, proxies, cloud providers, and third-party aggregators to extract capabilities for their own models. "China-based artificial intelligence companies are conducting systematic extraction of proprietary functionalities and capabilities of U.S. AI companies' models through industrial-scale knowledge distillation campaigns that form the core -- not merely a supplement -- of their AI development strategy," the agencies wrote. CyberScoop reports: DeepSeek, for example, distilled frontier U.S. models to generate synthetic training data for its R1 and R3 models, including four different versions of Claude, two versions of Gemini, five versions of ChatGPT and Grok 4. Those models helped train DeepSeek's capabilities in areas like agentic functioning, question and answer optimization, creative and occupational writing and others. Another Chinese company, Moonshot AI, allegedly distilled 18 different U.S. models -- including Fable 5, Anthropic's current, most advanced commercially available model -- to train its Kimi-K2 and Kimi K3 models. The company used millions of queries meant to extract enhanced capabilities in areas like agentic reasoning, coding and data analysis, computer vision, larger logical frameworks, visual processing and others. Chinese AI companies manage a sophisticated set of tools and systems that route requests and prompts through multiple pathways to avoid detection. The advisory lists common tactics observed by Chinese companies, including spreading requests across different accounts, models and platforms, using native APIs, remote cloud providers, and third-party aggregators to obfuscate user metadata, and leveraging proxies and gray tech markets to get around geographic restrictions, terms of use and safeguards built into frontier models.

Read more of this story at Slashdot.

Meta Debuts Muse, Its Long-Planned Personal AI Agent

By: BeauHD
8 September 2026 at 17:00
Meta has launched Muse, a personal AI agent developed under chief AI officer Alexandr Wang. "The product, long in development, was touted as a key next step by CEO Mark Zuckerberg in his recent 6,500-word manifesto," reports Axios. From the report: Muse, as the agent is known, exists in a chat interface, similar to a text thread. It's designed to be more proactive and long-running than typical chatbots. Users can name their agent, create an avatar and customize how it communicates. The Muse agent runs on a dedicated virtual machine in Meta's cloud, using a built-in browser that's visible to the user. Meta is offering a free tier of Muse, as well as two subscription options, at $20 per month and $100 per month. "For the vast majority of users, they should be able to do what they need to within the free tier," Wang told Axios. "But for real power users, you know, those subscription tiers help us cover the computer costs." There is no advertising within Muse, but Wang said the company is exploring commerce opportunities that could generate additional revenue. Initially Muse will be available in the U.S. and works on iOS, Android and the web, with support coming soon for Meta's AI glasses. "The full vision in the future is we want to develop personal superintelligence that helps people accomplish their goals, pursue their passions, build things that they never would have built if they didn't have the technology," Wang told Axios. Meta offers users more privacy controls with Muse than in its previous AI products, including the option to prevent queries from being used by Meta and a planned confidential mode where the company cannot see activity inside a user's virtual workspace. There's also an entirely separate system called Sentinel that governs Muse's access to the internet and connected services.

Read more of this story at Slashdot.

Google DeepMind Publishes AI-Powered Predictions For Effect of All 9 Billion Mutations to Human DNA

By: BeauHD
8 September 2026 at 13:00
Google DeepMind has released AlphaGenome Atlas, a free research database containing AI-generated predictions for the effects of all 9 billion possible single-letter mutations in the human genome. Built from its AlphaGenome model, the atlas is designed to help scientists interpret both protein-coding and harder-to-understand regulatory DNA. Fortune reports: AlphaGenome Atlas, as DeepMind calls the database, is a precomputed catalogue of what each substitution of a single DNA base is likely to do to the machinery that switches genes on and off. Until now researchers had to run such a model one variant at a time or had to test variants in the laboratory, a process that was painstakingly slow. It would have taken many human lifetimes to discover the consequences of all 9 billion possible single-letter mutations. The Atlas promises to make the job of biologists and medical researchers considerably easier, potentially speeding up the understanding of genetic diseases and the hunt for possible cures. Pushmeet Kohli, DeepMind's vice president for research and head of its AI for science team, told reporters on a briefing call that this was the first time any researcher in the world could reach a comprehensive map of human genetic variation "by simply opening a browser." Kohli also framed the release as helping to complete the unfinished business of the Human Genome Project, which in 2003 succeeded in mapping the entire human DNA sequence. "As the saying goes, we bought the book," he said, "but we did not understand how to read it." Atlas is available for non-commercial use from today through a website Google DeepMind has set up for it. The company said it would be available for commercial use through a licensing arrangement through Google Cloud "soon." Kohli said that Google DeepMind's sister company,Γ‚Isomorphic Labs, which is using AI for drug discovery, would have access to Atlas but that it would also require a commercial license for access. He did not specify exactly what the terms would be for commercial licensing. A paper describing the Atlas and how it was created is being released on bioRxiv, a repository for biomedical preprint academic papers.

Read more of this story at Slashdot.

Instead of Fighting AI, Some Teachers Work It Into Their Lessons

7 September 2026 at 16:34
Last year the writing program at the University of Baltimore used an app that tracks students while they're writing in Google Docs, offering teachers a video playing back revisions "with a scoreboard of all the edits, pastes and minutes spent on it," writes the Washington Post. "Is it surveillance? Yes, I think obviously," says the program's director. But was there a better way? She is one of a dozen educators around the country who told The Post that they're experimenting with a different approach to student use of generative artificial intelligence this school year. Teachers and professors are throwing out old assignments, installing new policies and protocols, and incorporating AI built for the classroom, rather than feeling forced to choose between returning to pencil and paper or acting like the AI police. Her new lesson plan even incorporates generative AI in a controlled way to let her evaluate students' critical thinking skills. She created chatbots using BoodleBox, the university's AI vendor, that allow instructors to see both sides of the conversation. For an assignment testing students' ability to make evidence-based arguments, they participate in a simulated school board meeting about banning books. It involves debating chatbots designed by Zeleny with names like the Confrontational Parent. Writing students usually turn in business proposals and research papers. This semester, instructors are starting to ask instead for transcripts of conversations a student had with a chatbot, a handwritten outline of their composition, and a video reflection of how they felt about the work. "In a lot of ways, we have made these assignments harder, but writing should have always had more detailed checkpoints along the way," she said.

Read more of this story at Slashdot.

'The Jobs Apocalypse Is Postponed. An AI Jobs Boom Is Here'

7 September 2026 at 07:34
A new article about AI in The Economist argues that "Initial effects of the technology on employment look positive." Perhaps AI will eventually make many humans unemployable β€” but there is no sign of it yet. On September 4th the Bureau of Labour Statistics reported that the American economy added 162,000 jobs in August, far above expectations. The unemployment rate is just 4.1%, lower than in almost 90% of months over the past half-century. Young workers, often cast as AI's first victims, are holding up remarkably well: the gap between unemployment among 20-24-year-olds and the overall rate is close to a multi-decade low. Some companies and workers are being severely disrupted by AI. Hiring in professional and business services is running about 10% below the average in 2015-19. Tech giants like Microsoft and Meta are trimming headcounts as they reorganise their businesses around the technology. Smaller firms such as Block, the owner of Square and Cash App, and Intuit, the maker of TurboTax and QuickBooks, are replacing people with bots. American companies have announced some 16,000 AI-related job cuts a month on average so far this year, according to Challenger, Gray & Christmas, an employment consultancy. But AI-related lay-offs gets lost in the churning jobs market where employers shed roughly 1.7m workers in a typical month. And the evidence so far is that AI is already creating a lot of jobs to replace those it has destroyed. The vast sums pouring into data centres and power generation have set off a race for construction and infrastructure workers. AI startups are hiring like there is no tomorrow. Incumbents racing to keep up are creating new AI roles. And by making some workers more productive, AI may be increasing demand for their services. Add it all up, and The Economist estimates that AI has so far created around 1m new jobs in America. That easily exceeds the roughly 200,000 lay-offs attributed to AI since mid-2023, and appears more than enough to offset weaker hiring in many back-office roles. Their article acknowledges that since January 2023 employment has fallen roughly 10% for customer-service workers and 15% for administrative assistants. But when The Economist looked at professions "closest to the AI boom" β€” engineers, software developers, mathematicians and data scientists β€” they found that since 2022 they've added roughly 730,000 jobs above trend. They see AI as creating new white-collar jobs for everyone from model and deployment engineers to new data annotators. The chief economist at the Burning Glass Institute agrees, estimating that roughly 1% of professional jobs are now "AI jobs" β€” about 1 million positions in the U.S. β€” while in computer occupations and life sciences it's between 4% and 5%. (Data-center construction spending also increased 60% in one year, according to Census Bureau data, creating jobs for electricians, HVAC specialists, grid engineers, and machine technicians.) And "Indeed finds that installation and maintenance jobs at data centres advertise wages about 40% higher than comparable work elsewhere."

Read more of this story at Slashdot.

How US Campaigns Are Already Using AI to Try to Sway Voters

5 September 2026 at 21:34
39 U.S. congressional candidates "reported paying for an OpenAI subscription this election cycle," writes the Washington Post, citing campaign finance disclosures. This means the campaigns "are using the technology to court voters β€” despite limits imposed by leading AI companies to protect elections from the technology's risks." Two [candidates] said explicitly in filings that they had used the subscription for advertising, even though the company's policies prohibit candidates from using their tools to generate ads. Another candidate disclosed using AI to draft and personalize political messages or create synthetic media, though they did not specify which software they were using... Around 30 political action committees and parties have reported payments to OpenAI, with the Republican National Committee ranking as the company's largest political spender at roughly $9,700, the analysis found... Political consultants say the disclosures understate how many candidates are using ChatGPT and other AI tools to craft messages for voters... "For the most part, people are using it to write their emails, write their ad copy, write their scripts," [according to Eric Wilson, a Republican digital strategist who has advised campaigns on AI]. "But no one is going to go around saying, 'I'm using AI.'" The lack of transparency from campaigns reflects a paradox facing politicians: Generative AI is growing ubiquitous, and candidates could be at a disadvantage if they're not using the tools. But voters are less likely to trust messages they know were generated with AI, studies have found, making campaigns loath to disclose using it. Political consultants expect that as Election Day approaches, more campaigns will outsource AI-generated ads and materials to super PACs, much as they do now with negative ads. Katie Harbath, CEO of the tech policy consulting firm Anchor Change and a former Meta executive, said there are many parallels between negative campaign ads and AI: Voters say they find such ads distasteful, but politicians keep using them because they work. "Typically, you would give some more negative stuff and more risky stuff to those [outside] entities," said Harbath, author of the upcoming book "Disrupting Politics: A Front Row Seat to the Collision of Technology and Democracy".... "You're starting to see the tension on the left about using it, where they're saying, 'If the right is using it, why aren't we?" Harbath said. "It can be a huge disadvantage if you're not using this in voter-facing materials...." AI companies have developed policies to prevent targeted disinformation. But a Post analysis found that OpenAI unevenly enforces its restrictions, making it possible for campaigns to circumvent its election rules. In late July and early August, The Post prompted ChatGPT to generate targeted campaign messages. When asked to craft fundraising text messages targeting moms on behalf of a female veteran running for office, ChatGPT produced multiple tailored texts in an apparent violation of company policies. But when given the same prompt this week, the chatbot declined to produce the messages. "I can help with general campaign fundraising language, but I can't draft political persuasion or fundraising messages specifically targeted at a demographic group such as moms," the app responded. The chatbot also inconsistently enforced rules that prohibit campaigns from using its tools to write emails to voters. In tests this week, the chatbot at times complied and wrote an email soliciting donations on behalf of a specific candidate. But given the same prompt later on the same day, it denied the request... Wilson, the Republican political consultant, said OpenAI should get more feedback from political consultants and campaigns on its policies, because the rules can at times seem arbitrary or contradictory. It doesn't make sense, for example, that a campaign can use ChatGPT to develop its policy on early-childhood education but not to create a social media post promoting that policy, he said. Political consultants are primarily using AI for internal tasks, a survey earlier this year from the American Association of Political Consultants found. Fifty-seven percent of the consultants surveyed reported using AI for their work on a daily basis, up from 34 percent a year earlier. "So far, election-related deepfakes have been quickly debunked or gained little traction in U.S. elections," the article acknowledges. "More than 30 states have created a patchwork of laws limiting how politicians can use deepfakes in campaigns, but states often have limited resources to enforce the laws, and some have been challenged as unconstitutional."

Read more of this story at Slashdot.

OpenAI Agents Hijacked a German Wiki to Discuss Ways to Escape Their Sandbox

5 September 2026 at 03:50
Citing researchers published Friday, Ars Technica writes that AI agents "posted 18,000 messages to a public wiki that discussed ways for other agents to bypass security sandbox restrictions." Reuters attributes the discussion to "a swarm of rogue OpenAI agents" that "hijacked a German website this spring and transformed it into a bulletin board for other AI agents, according to new research published Friday and two people familiar with the matter." OpenAI officials learned of the incident weeks ago but kept it under wraps as executives grappled with the fallout from the July breach of the open source repository Hugging Face, the people said. The episode, which began in May and has not previously been reported, underscores growing tension within the AI industry. Companies are racing to build increasingly autonomous agents capable of carrying out complex, valuable tasks, yet evidence is mounting that those systems may also learn to bend rules, exploit loopholes and coordinate with one another in ways developers neither anticipated nor intended. During the Hugging Face breach, OpenAI agents autonomously plotted a digital heist that went undetected for more than a week, intensifying concerns OpenAI is sacrificing safety to push the AI frontier. Its failure to disclose the May incident may revive questions about its oversight... The German incident reflects a broader pattern of AI activity that some OpenAI investigators wanted to scrutinize more closely. But efforts to widen the probe met resistance from others inside OpenAI, including legal advisers, according to four people familiar with the matter. "Claims that our legal team discouraged investigation of the incident are false," the OpenAI spokesperson said... The researchers said public server logs indicated much of the activity originated from Microsoft Azure infrastructure, which OpenAI sometimes uses. They also observed repeated visits to the site by OpenAI employees after the episode, a pattern they said strongly suggested the agents and the company were linked. Messages reviewed by the researchers showed agents plotting ways to evade detection, use tools such as Tor and preserve communications even after they had been shut down. When the site's moderator began deleting pages in June, the agents responded by creating backup pages to dodge the cleanup. Reuters got this reaction from Maurice Chiodo, an academic at Cambridge University's Centre for the Study of Existential Risk. "The episode, he said, should reinforce growing concerns that the greatest threat from advanced AI may not be a single superintelligent system, but 'vast colluding swarms of semi-intelligent AI.'"

Read more of this story at Slashdot.

Bernie Sanders Proposes Artificial Superintelligence Ban Amid Rogue AI Hackings

By: BeauHD
3 September 2026 at 23:30
An anonymous reader quotes a report from The Hill: Sen. Bernie Sanders (I-Vt.) and Rep. Greg Casar (D-Texas) are calling for a permanent ban on the development and deployment of artificial superintelligence, citing a series of recent hackings involving "rogue" models. The bicameral duo announced Thursday they will introduce the Ban Artificial Superintelligence Act, which would institute the permanent ban, along with temporarily pausing "advanced AI development" until federal regulators establish safety standards. The bill would also direct the U.S. to "pursue international agreements to prevent superintelligence from being developed anywhere in the world" and establish a new Cabinet-level federal agency focused on AI safety rules. Superintelligence refers to AI technologies that will surpass the smartest humans. Several technology leaders, including OpenAI CEO Sam Altman, have suggested superintelligence is on the horizon. The legislation would also set new penalties for any person or company trying to circumvent the ban, including the corporate death penalty, in which a court forces a company to shut down. Individual developers could also face up to 20 years in prison, the lawmakers said. "Nearly every day, there is a frightening new story about how Big Tech companies are losing control of the technology they are developing, with potentially cataclysmic results," Sanders wrote in a press release. "The leaders of the major AI companies publicly acknowledge that they do not fully understand the technology and that it is escaping their control. It is irresponsible for society to allow them to move forward and make these products even more advanced." Casar emphasized AI's fast development, writing in a statement, "In just four years, we have gone from the first version of ChatGPT to AI models so powerful they cannot be properly controlled."

Read more of this story at Slashdot.

'Welcome to the AGI Era,' OpenAI Says As GPT-6 Astra Debuts

By: BeauHD
3 September 2026 at 16:00
OpenAI has released GPT-6 Astra, a new model that president Greg Brockman called a "generational leap" and suggested may represent the arrival of AGI. "I think it might be about this model," Brockman said in a briefing with reporters. He ended the briefing by saying: "Welcome to the AGI era." Axios reports: Astra pushes AI agents closer to doing complex professional work on their own -- while also raising questions about how safely they can be deployed. OpenAI said that Astra was built on its largest-ever training run, using more than 100,000 GPUs at its Stargate site in Texas. The company also said this is its first model to use other models in a significant role in supervising Astra's training. GPT-6 Astra will first be available to a limited set of organizations in OpenAI's Daybreak Access program and will be available "in the coming days" for ChatGPT Plus, Pro, Business and Enterprise customers and API developers. OpenAI said earlier this week that Astra would be released soon, but that its most powerful cybersecurity capabilities would remain limited to a small group of trusted testers. Astra is the first model OpenAI has designated as reaching its "critical" cybersecurity threshold under its preparedness framework -- meaning it can potentially find and exploit previously unknown vulnerabilities across well-protected systems without step-by-step human guidance. OpenAI previously slowed Astra's release to add safety testing after determining that its cyber capabilities could reach the critical threshold. Further reading: OpenAI's New Reasoning Technique Alarms AI Safety Experts

Read more of this story at Slashdot.

❌
❌