Showing posts with label OpenAI. Show all posts
Showing posts with label OpenAI. Show all posts

Thursday, October 08, 2026

Sharing AI progress in mathematics by released more than 700 AI-generated papers | OpenAI

This is only the beginning! What an avalanche! Breathtaking!

How fast will mathematics (the queen of the sciences according to Carl Friedrich Gauss) now advance going forward? Maybe in the coming months/quarters we will see more progress than in the past 1000 years or so.

"OpenAI released more than 700 AI-generated papers presenting results related to more than 370 previously unsolved mathematical problems, including some of the field’s hardest. The massive dump marks yet another landmark in the use of AI to attack complex math—but experts are split over whether such releases are good for the field. Learn about what’s in the papers—and how researchers are reacting."

Sharing AI progress in mathematics | OpenAI

OpenAI’s $25 billion research charity

Good news!

"The OpenAI Foundation — the charity arm of the artificial-intelligence behemoth — is on track to become one of the wealthiest non-profit organizations on the planet.
It has so far provided a $40 million grant for a project at the University of North Carolina at Chapel Hill to develop cancer vaccines, as well as funding for Alzheimer’s disease research and a pilot project to repurpose data from bankrupt biotechnology companies. ..."

Nature Briefing: Cancer

How to spend $25 billion on science: OpenAI’s staggeringly rich research charity gets under way "Jacob Trefethen, who leads life sciences at the OpenAI Foundation, speaks to Nature about the organization’s ambition to cure disease."

OpenAI Foundation "Our mission is to ensure artificial general intelligence benefits all of humanity."

Supporting Communities: Meet the 2026 People-First AI Fund Grantees "Our second People-First AI Fund is deploying $50 million to 163 nonprofits exploring how AI can help communities across the United States access essential services, strengthen creative and cultural institutions, and support local journalism."


Every grant begins with people. These are some of the communities the People-First AI Fund exists to support.


Thursday, September 24, 2026

Australia to investigate if OpenAI hack of government health website broke the law

Possible trouble for OpenAI from down under? We are learning more about how OpenAI may have obtained data.

"An OpenAI model hacked into an Australian government website, the country’s prime minister Anthony Albanese said Wednesday, in the first publicly reported case of an AI model hacking into a government’s systems. 

Albanese said that there would “obviously be legal consequences” following the breach, and that OpenAI faces a government investigation into how its unreleased models gained access to reams of bulk health data information. ..."

Australia to investigate if OpenAI hack of government health website broke the law | TechCrunch

Monday, September 14, 2026

Build more natural voice experiences with full-duplex GPT‑Live‑1 in the API

Good news!

"OpenAI releases full duplex voice model to the API

OpenAI released GPT-Live-1, a voice model that listens and speaks simultaneously rather than chaining separate speech-to-text, reasoning, and text-to-speech models.
It handles interruptions and background noise within a single model, delegating deeper reasoning and tool calls to a backend such as GPT-6 Astra, and supports telephony deployments for phone-based customer support. 
OpenAI reports a 30-percentage-point improvement on Full Duplex Bench over GPT-Realtime-2.1 and first place on Tau3 when paired with GPT-6 Astra;
Speak’s early evaluation showed an 80 percent cut in interruptions versus turn-based systems. Pricing is $0.05 per minute for the voice layer, with backend model and tool costs billed separately; it natively supports ASR transcripts, keyword biasing, and turn detection. For voice app developers, it removes the latency and coordination overhead of stitching together separate STT, LLM, and TTS components." (Data Points)

Build more natural voice experiences with GPT‑Live‑1 in the API | OpenAI "GPT‑Live‑1 brings ChatGPT’s natural, full-duplex conversations to the API, with more control over how voice agents speak and act."

Tuesday, September 08, 2026

On the Navier–Stokes Millennium Prize Problem | OpenAI

Amazing stuff! Exciting times for math!

However, OpenAI fought dirty on career-making math problem, says NYU mathematician. "OpenAI fought dirty on career-making math problem, says NYU mathematician: There is a $1 million bounty for the first person providing a solution to the Navier-Stokes existence and smoothness problem." OpenAI admits there was apparently an intense competition going in the last few days or so on who would come out first.

"We’re sharing a solution to the Navier–Stokes existence and smoothness problem, one of the Millennium Prize Problems. This proof, produced by an internal OpenAI system, shows that the dynamics of the Navier-Stokes equations for fluid motion can develop a singularity in finite time. We’re sharing both a writeup of the proof and a formalization in Lean. ...

The agents arrived at their resolution on Saturday, September 5, about 88 hours after the first agents were launched. Lean formalization and verification took an additional 17 hours via GPT‑6 Astra. ...

In the process of resolving the Navier–Stokes problem, the agents sent 2.7 million messages and used approximately 130 billion output tokens. ..."

On the Navier–Stokes Millennium Prize Problem | OpenAI


A snapshot of local incompressible motion. Orange marks faster angular rotation; teal marks slower rotation. Circulating speed also depends on radius. The trajectories show inward spiraling and axial stretching.


Thursday, September 03, 2026

A milestone in expanding access to AI (via advertising-supported free tier) | OpenAI

Good news!

Unfortunately, this article does not explain what the advertising-supported free tier is and how to access it?

"In less than 200 days after launch, ChatGPT Ads has reached $1 billion in annualized revenue run rate. The platform is now used by tens of thousands of advertisers and continues to expand globally.
Starting later today [8/31/2026], advertisers can purchase ChatGPT ads directly via Ads Manager across India, Europe, the Middle East, and North Africa.

Advertising is one pillar of OpenAI’s diversified business model, alongside consumer subscriptions, enterprise offerings, and usage-based APIs.
Together, these offerings give people, developers, and businesses choice in how they access OpenAI products, including an advertising-supported free tier that helps keep ChatGPT available to more than 1 billion weekly active users. ...

Ads in ChatGPT are always clearly labeled and separate from ChatGPT’s answers. Advertising does not influence the answers ChatGPT provides. Advertisers do not receive access to people’s private conversations, and users can control how their ad experience is personalized. ...

ChatGPT Ads are now available in over 40 countries⁠ ... through the OpenAI Ads Solutions team and our agency and technology partners. Starting later today, self-service access to ChatGPT Ads is launching across India, Europe, the Middle East, and North Africa."

A milestone in expanding access to AI | OpenAI "ChatGPT Ads reaches $1 billion in annualized revenue run rate and global expansion continues."

Anthropic sued (again) for training on song lyrics and US government sides with OpenAI on issue of training LLMs on copyrighted material

What is the training of large models on copyrighted material (intellectual property)? Fair use or pirating? This answer to this question becomes very urgent.

The statutory damages asked for by plaintiffs could easily bankrupt Anthropic or any other company involved in model training.

I again strongly suggest to reform intellectual property rights. They are effective for way too many years! E.g. books and songs in the US, they "last for the life of the author plus an additional 70 years." (Google search) This is insane!!!

"In a lawsuit that The New York Times filed against OpenAI, the Trump administration has contributed a 20-page brief in defense of the ChatGPT maker’s unlicensed use of copyrighted material to train its LLMs.

“The United States has a strong interest in continuing to develop a robust and competitive artificial intelligence industry that sets the standard for the practice and procedure of AI use globally… As such, it is critical for the United States to ‘retain global leadership in artificial intelligence,’” the brief reads, referencing an executive order that President Donald Trump signed last year. ..."

"Sony Music Publishing and Warner Chappell Music sued Anthropic, along with CEO and co-founder Dario Amodei and co-founder Benjamin Mann, in the U.S. District Court for the Northern District of California, alleging Claude was trained on tens of thousands of copyrighted songs without permission. The complaint accuses Anthropic of torrenting pirated books from Library Genesis and Pirate Library Mirror, scraping lyrics from licensed sites like Musixmatch and LyricFind, and running a “destructive scanning” operation on physical books, citing findings already unsealed in the separate Bartz v. Anthropic authors’ case. The publishers seek statutory damages up to $150,000 per willfully infringed work plus up to $25,000 per removed copyright notice, destruction of infringing copies, and disclosure of Claude’s training data.
This is the fifth major music-industry suit against Anthropic, following actions from Universal Music Publishing Group, Concord, and ABKCO (over $3 billion), BMG, and Round Hill Music.
It also follows Anthropic’s $1.5 billion settlement with book authors over similar torrenting conduct, announced in 2025 and approved this year. The suit signals that music publishers are pushing for damages and licensing terms rather than accepting a one-time settlement. ..." (DataPoints)

Now Sony Music Publishing and Warner Chappell sue Anthropic in multi-billion dollar lawsuit: ‘One of the largest and most blatant ongoing thefts of intellectual property in history.’

US government sides with OpenAI on issue of training LLMs on copyrighted material | TechCrunch

Tuesday, August 25, 2026

OpenAI’s Jalapeno chip is built for fast inference at scale, benchmarks show

Good news! OpenAI is also getting into the chip business.

Why did OpenAI opt to write Jalapeño with the Spanish special character instead of a clean and simple Jalapeno? Grrr! This special character is not on my keyboard! I hope, OpenAI will correct this mistake! Keep it simple stupid (KISS)!

"At the Hot Chips conference on Tuesday [8/25/2026], OpenAI shared a more detailed look at Jalapeno, including the first batch of benchmark results for the new system. Tested on SemiAnalysis’ InferenceX benchmark, Jalapeño registered both more tokens per user and more throughput per kilowatt than the currently available state-of-the-art inference processors. ...

First announced last October, Jalapeño was developed by OpenAI in close collaboration with Broadcom, with OpenAI’s own models assisting in the development process. The company plans to make Jalapeño a multigenerational platform, allowing AI products, models, chips, and memory all developed in concert.

Because of that full-stack approach, OpenAI was able to address specific phases in the inference process that often cause friction during inference processing. In particular, Jalapeño is designed to minimize delays during the prefill and communication phases of processing, which OpenAI says often act as bottlenecks.

“We designed Jalapeno to minimize data movement and communication delays,” ..."

"... OpenAI models also accelerated Jalapeno’s development. Earlier generations helped the team design and bring up the chip, while our latest models are accelerating how we optimize and program it. Jalapeno’s performance extends across GPT‑OSS 120B, DeepSeek R1, and Kimi K2.5 1T, showing that the architecture works across models developed both inside and outside OpenAI.
Across all three, Jalapeno delivered 1.5 to 1.9 times more AI work per watt at peak throughput and 1.7 to 3.6 times lower end-to-end latency than the comparison systems. For highly interactive workloads, it delivered 2.1 to 4.1 times higher performance. ..."

OpenAI’s Jalapeno chip is built for fast inference at scale, benchmarks show | TechCrunch







Tuesday, August 04, 2026

Ten advances in mathematics and theoretical computer science | OpenAI

Amazing stuff! This is only the beginning!

"OpenAI released solutions to ten open problems in mathematics and theoretical computer science, generated by Astra, an internal unreleased model.
The problems span sphere packing, coding theory, group theory, and quantum complexity—each vetted by formal proofs in Lean.
Generating all ten solutions cost roughly $2,000 in compute at current API rates. The results include a disproof of Connes’s rigidity conjecture, new bounds on sphere-packing density, and a construction proving the existence of non-sofic groups. OpenAI framed the release carefully, noting that humans prepared the manuscripts but the mathematical arguments came from the system, a deliberate stance on authorship and attribution as AI systems edge into research collaboration."

From the abstract:
"We present a collection of results obtained by an internal OpenAI model, spanning mathematics and theoretical computer science:
1. High-dimensional sphere packing. The asymptotic strength of the Cohn–Elkies linear program is determined exactly. This gives an improved general packing bound in high dimensions and settles the corresponding Fourier sign-uncertainty problem asymptotically.
2. Binary and spherical codes. Classical upper bounds for fixed-distance binary and spherical codes are improved by exponential factors for all parameters. The spherical construction also recovers the sphere-packing exponent of Chapter 1.
3. Non-sofic groups. An explicit non-sofic group is constructed, resolving the question of whether every countable group admits finite permutation approximations. The argument uses property-(T) expanders and the binary Leavitt algebra.
4. Connes’s rigidity conjecture. Infinitely many pairwise nonisomorphic property-(T) groups are constructed with the same group von Neumann algebra, disproving Connes’s conjecture and answering a related finite-to-one question of Popa.
5. Arithmetic circuit complexity. For the permanent, division-free circuits require Ω(n2 log log n) gates, while formulas require Ω(n 4/ log n) leaves.
6. Quantum parallel repetition. Exponential parallel repetition is proved for every finite two-player entangled game, extending the classical repetition principle beyond previously treated special classes of quantum games.
7. Closest vector problem. A direct reduction from 3SAT gives n
1/400-factor hardness for the Euclidean closest vector problem, with related consequences for binary decoding and other lattice norms.
8. Ehrhart’s volume conjecture. The sharp bound (n + 1)n/n! is proved in every dimension for convex bodies whose barycenter is their only interior lattice point. 9. Multicolor Ramsey numbers. A superexponential lower bound proves Rk(3) = kΘ(k).
10. Compactness and degeneracy. Separate bipartite graph constructions disprove two conjectures in extremal graph theory: the compactness conjecture of Erdős and Simonovits and a degeneracy conjecture of Erdős."


Ten advances in mathematics and theoretical computer science | OpenAI

Ten Advances in Mathematics and Theoretical Computer Science (open access, not peer reviewed, 249 pages)

Monday, July 27, 2026

A short biography of one of the best known engineers of machine learning & AI: Ilya Sutskever

Recommendable!

(831) The Greatest AI Engineer of All Time - YouTube


OpenAI Launches Health Platform to the public at large in the US

Good news! Self medication and self treatment most likely preceded the emergence of the medical profession in human history!

Physicians/medical doctors are sometimes referred to in German language as "Halbgötter in Weiß" (demigods in white clothing)! Iconoclasm by AI?

Democratize healthcare! Pronto! 😊

Unfortunately, it appears that the first release is too much focused on Apple Health only. Hopefully, this will quickly be corrected by OpenAI.

"OpenAI rolled out ChatGPT Health, a service that allows users to securely connect their medical records and wellness apps to the artificial intelligence chatbot. 

The new feature comes in response to nearly 300 million people who already turn to ChatGPT with health-related questions each week. But users are cautioned to consider the likelihood of inaccuracy as well as the issue of sharing sensitive, potentially deeply personal information with a chatbot. ..."

The Flyover - Daily Newsletter for You

Launching Health in ChatGPT (original news release) "Rolling out to U.S. users, you can securely connect your health information to help you better understand and navigate your health."




'Unprecedented' Self-Directed CyberAttack on Hugging Face by OpenAI's Rogue Model Was Repelled by Chinese Defensive AI

I had intended to publish this blog post earlier! I procrastinated! Mea culpa!

I have just blogged here about this complex and very sophisticated cyber attack.

'Unprecedented' Self-Directed CyberAttack by OpenAI's Rogue Model Was Repelled by Chinese Defensive AI "AI startup Hugging Face deployed a Chinese-developed AI model to counter an unprecedented cyberattack launched by a rogue OpenAI system, highlighting both the emerging threat of autonomous AI attacks and the growing capabilities of Chinese AI technology."

OpenAI used the ExploitGym to cyber attack Hugging Face

What an irony! 😊 There is even a research paper published in May as a preprint on arXiv about ExploitGym with the title containing "Turn Security Vulnerabilities into Real Attacks"! 

As they say reality is stranger than fiction! 😊

I blogged here and here about this incidence before.

"... OpenAI pitted its models against a benchmark called ExploitGym, released in May, which challenges LLMs to find ways to exploit real-world vulnerabilities found in commonly used software. ...

On July 9, according to reporting by Reuters, OpenAI’s models started trying to break through the proxy. They found an unknown bug in the proxy’s software and used it to access the internet. From there, they broke into Hugging Face’s computer systems on July 11, apparently looking for data sets and solutions that would help them complete their task. Hugging Face announced the hack on July 16. ..."

The Hugging Face AI hack is scary—but unsurprising

Tuesday, July 21, 2026

The cyber attack on Hugging Face was quite complex and sophisticated


Bad news!

"An autonomous AI agent breached Hugging Face’s production infrastructure through a malicious dataset, moving laterally across systems for an entire weekend without detection.
When the incident response team tried to analyze the attack using commercial frontier models, safety guardrails blocked every forensic query—treating the defenders’ real exploit data the same way they would treat a live attack.
The agent executed thousands of actions through short-lived sandboxes, harvesting cloud credentials and reaching multiple internal clusters, all without human guidance.
Hugging Face ultimately completed its forensic analysis using GLM 5.2, an open-weight model running on its own infrastructure, because it was the only option that wouldn’t refuse to process attacker artifacts.
The incident exposes a basic asymmetry: Defenders operating under enterprise governance hit safety controls that don’t constrain attackers running uncensored models, turning AI tooling into a potential single point of failure during the exact moment security teams need it most."

"... The entry point was a malicious dataset. When the data-processing pipeline ingested it, the dataset triggered code execution in two ways, through a remote-code loader and through a template-injection flaw in dataset configuration files.
No admission gate screened the file before it reached a processing worker, the compute job assigned to each dataset.
Initial access came through the data pipeline itself, a route that worked because most enterprise security teams treat the data feeding their pipelines as trusted input rather than an attack surface.

Worker isolation did not hold. From that first foothold, the agent broke out onto the node running the workload, then harvested cloud and cluster credentials scoped broadly enough to reach multiple internal clusters, all of it inside a single weekend. Hugging Face described the agent framework as executing thousands of individual actions through a swarm of short-lived sandboxes, with self-migrating command-and-control staged on public services. The disclosure adds that the framework appeared to be built on an agentic security-research harness, which would put tooling designed for red-team work behind a live intrusion. ..."

Data Points: Kimi K3 marks a big shift in AI development

Thursday, July 09, 2026

How China Is Meddling in America’s AI Debate by covert operations discovered by OpenAI

Like the heavy meddling of the former Soviet Union, I have no doubt that this is happening! The war of words and propaganda wars!

By crippling Western AI advances in the court of public opinion, China is trying to get ahead!

It is truly astonishing how little attention this OpenAI report (released 6/10) got in the news media! Are they already too scared or too much under the control of the Communist Party of China?

"... We now know that another extraordinarily influential force has also put its thumb on this crucial American policy discussion: the People’s Republic of China.

In a blockbuster report issued last week, entitled “PRC-linked influence operations are targeting AI debates in the US,” tech giant OpenAI found that actors originating in China “used our models in support of apparent covert influence operations that promoted narratives in an attempt to manipulate a legitimate debate about American AI and wider tech policies.”
Specifically, one cluster of Chinese users of OpenAI’s platform, in plain violation of its terms and conditions, “generated social media comments and images claiming that data center buildouts for AI were increasing electricity prices for average families.”
A second cluster produced and disseminated content “criticizing US tariffs as attempts to dominate technological competition and specified in their prompts that the content should not include China’s leader Xi Jinping in the output and instead include only President [Donald] Trump.” ..."

"... The accounts we [OpenAI] banned sought to influence two groups of audiences. They primarily targeted US audiences and generated English-language short comments and images claiming that data centers and AI applications were increasing electricity demand and causing higher costs for ordinary Americans.  ..."

How China Is Meddling in America’s AI Debate | American Enterprise Institute - AEI






Monday, May 25, 2026

OpenAI model finds proof resolving famous mathematics problem dating from 1946

Amazing stuff! More to come! This is only the beginning!

"OpenAI model finds proof resolving famous mathematics problem

An OpenAI reasoning model has resolved the planar unit distance problem, a central question in discrete geometry posed by legendary mathematician Paul Erdős in 1946.
The conjecture held that square grid constructions were essentially optimal for maximizing unit-distance pairs among points in a plane, a belief that stood unchallenged for nearly 80 years.
The model instead found an infinite family of configurations yielding polynomial improvements over the grid approach, disproving the assumption.
What makes the breakthrough unusual is not just the result itself, but how it was found: a general-purpose reasoning model, not a system specialized for mathematics, produced a proof that external mathematicians have verified. The proof brings sophisticated tools from algebraic number theory to bear on an elementary geometric question, revealing unexpected connections between distant mathematical domains. Fields medalist Tim Gowers called it “a milestone in AI mathematics,” while number theorist Arul Shankar argued the result shows AI models “are capable of having original ingenious ideas, and then carrying them out to fruition.”" (Source)



Paul Erdos (Source)


Previously known construction of many unit distances from a rescaled square grid.


Monday, May 04, 2026

AI is starting to beat human doctors at making correct diagnoses in the emergency room

Good news! Finally, human doctors/phycisians get some serious competition to the benefit of our health!

"... This vision of AI-assisted emergency health care may soon be reality. In a new study, researchers show that a type of AI known as a large language model (LLM) often outperformed physicians at diagnosing complex and potentially life-threatening conditions, including decreased blood flow to the heart, even in the fast-moving stages of real ER care when information is limited, they report today in Science.
In early ER cases, the model identified the correct or a very close diagnosis in about 67% of cases, compared with roughly 50% to 55% for physicians. And the technology is only getting better. ..."

From the editor's summary and abstract:
"Editor’s summary
Computational tools for medical decision support have been advancing over time, mainly by serving as resources for limited applications. Machine learning tools for autonomous interpretation of clinical cases have also been gradually improving over time.
Brodeur et al. pitted a large language model, the OpenAI o1 series, directly against hundreds of physicians at different levels of training and experience on a variety of clinical cases ranging from published patient vignettes to evaluations of brand-new emergency room patients, as well as on clinical tasks including both diagnosis and planning of clinical management  ... Across a variety of scenarios and applications, the large language model outperformed both human physicians and older models, suggesting its potential utility for clinical care.  ...

Abstract
More than 65 years ago, complex clinical diagnostic reasoning cases were introduced as the gold standard for the evaluation of expert medical computing systems, a standard that has held ever since.
In this study, we report the results of a physician evaluation of a large language model (LLM) on challenging clinical cases across five experiments with a baseline of hundreds of physicians.
We then report a real-world study comparing human expert and artificial intelligence (AI) second opinions in randomly selected patients in the emergency room of a major tertiary academic medical center.
In all experiments, the LLM outperformed physician baselines and displayed continued improvement from prior generations of AI clinical decision support. Our study suggests that LLMs have eclipsed most benchmarks of clinical reasoning, motivating the urgent need for prospective trials."

AI is starting to beat doctors at making correct diagnoses | Science | AAAS "Large language model excels at clinical decisions, even in fast pace of a simulated ER"

AI can reason like a physician—what comes next? (Perspective, open access) "Text-based AI can think like a physician; the challenge is achieving safe clinical implementation" [This Perspective has no abstract!]

Sunday, February 22, 2026

Sam Altman would like to remind you that humans use a lot of energy (and water), too

Good point! Touché!

Let's also keep in mind there is a high probability that the global human population will shrink in the coming decades due to a lower global total fertility rate (TFR).

"... For one thing, Altman ... said concerns about AI’s water usage are “totally fake,” though he acknowledged it was a real issue when “we used to do evaporative cooling in data centers.” ..."

Sam Altman would like to remind you that humans use a lot of energy, too | TechCrunch


Sam Altman


Thursday, January 15, 2026

OpenAI signs deal, worth $10B, for compute from Cerebras

Good news! Chip competition is good, more chip competition is better!

"... OpenAI announced Wednesday [14/1/2026] that it had reached a multi-year agreement with AI chipmaker Cerebras. The chipmaker will deliver 750 megawatts of compute to the AI giant starting this year and continuing through the year 2028, Cerebras said. ..."

"OpenAI and Cerebras have signed a multi-year agreement to deploy 750 megawatts of Cerebras wafer-scale systems to serve OpenAI customers. This deployment will roll out in multiple stages beginning in 2026, making it the largest high-speed AI inference deployment in the world. ..."

"Cerebras builds purpose-built AI systems to accelerate long outputs from AI models. Its unique speed comes from putting massive compute, memory, and bandwidth together on a single giant chip and eliminating the bottlenecks that slow inference on conventional hardware. ..."

OpenAI signs deal, worth $10B, for compute from Cerebras | TechCrunch


OpenAI partners with Cerebras (original news release) "OpenAI is partnering with Cerebras to add 750MW of ultra low-latency AI compute to our platform."

Credits: Last Week in AI


A gigantic chip indeed!