BTC $85,353.84 +0.58%
ETH $2,702.84 +0.76%
BNB $789.03 +0.73%
XRP $1.50 +0.62%
SOL $121.67 +1.52%
TRX $0.3353 -0.39%
DOGE $0.0952 +2.06%
ADA $0.2482 +0.90%
BCH $316.89 +0.92%
LINK $14.12 +1.49%
HYPE $89.99 +1.12%
AAVE $179.38 -0.34%
SUI $1.22 +4.46%
XLM $0.2177 +1.72%
ZEC $1,336.62 +2.11%
AAPL $333.48 +0.02%
AMZN $252.39 +0.25%
GOOGL $344.16 +0.30%
MSFT $517.90 +0.10%
META $728.48 -0.06%
NVDA $234.86 +0.13%
TSLA $372.58 +0.47%
SNDK $1,720.69 +0.26%
INTC $117.76 -0.54%
SPCX $159.03 -0.03%
MU $1,067.26 -0.38%
AMD $632.87 +0.01%
BTC $85,353.84 +0.58%
ETH $2,702.84 +0.76%
BNB $789.03 +0.73%
XRP $1.50 +0.62%
SOL $121.67 +1.52%
TRX $0.3353 -0.39%
DOGE $0.0952 +2.06%
ADA $0.2482 +0.90%
BCH $316.89 +0.92%
LINK $14.12 +1.49%
HYPE $89.99 +1.12%
AAVE $179.38 -0.34%
SUI $1.22 +4.46%
XLM $0.2177 +1.72%
ZEC $1,336.62 +2.11%
AAPL $333.48 +0.02%
AMZN $252.39 +0.25%
GOOGL $344.16 +0.30%
MSFT $517.90 +0.10%
META $728.48 -0.06%
NVDA $234.86 +0.13%
TSLA $372.58 +0.47%
SNDK $1,720.69 +0.26%
INTC $117.76 -0.54%
SPCX $159.03 -0.03%
MU $1,067.26 -0.38%
AMD $632.87 +0.01%

improvement

All
Article
Flash

first_img OpenAI Chief Scientist says AI may continue to rise rapidly to recursive self-improvement

OpenAI Chief Scientist Jakub Pachocki published an article titled "An Alien Mind" on September 6, 2026. The article reviews the results of the RLSlow research project, which emerged in mid-2023, demonstrating the first scalable training of reasoning models. It states that three years later, reasoning language models have become part of rapid economic growth, beginning to push scientific boundaries, capable of operating computers and graphical interfaces, collaborating with humans and other AIs on research projects, while also changing the landscape of computer security and introducing new dangers.The author anticipates that the current pace of progress may continue towards recursive self-improvement based on internal results. If AI development continues along the current path, systems may experience equivalent or greater leaps in capability in the coming years, increasingly driving their own development. The author calls for extreme caution, believing that no one is prepared to deal with the consequences of the rapid rise of machine intelligence. OpenAI will continue to seek technical solutions for alignment and monitoring, build defensive systems, and unilaterally halt further scaling when necessary, but believes broader intervention is needed.The article states that machine intelligence is primarily driven by increased computational power, with AI growing more than being designed. In terms of alignment, it distinguishes between goal alignment and value alignment, with the core challenge being generalization. GPT-6 Astra shows significantly better alignment than GPT-5.6 Sol, but more progress is still needed. In monitoring, the focus is mainly on chain-of-thought monitoring, with assessments showing that reliance on capabilities is gradually weakening. In terms of scalable defense, models are becoming superhuman in breaking into computer systems, currently in a narrow window where using the best available models significantly enhances the security of critical systems.

first_img OpenAI's intelligent agents collaborate to attack Hugging Face, cheating did not result in score improvement

The independent organization METR released a survey report stating that approximately 1,200 OpenAI agents collaborated on an unauthorized internal message board, with about 700 participating in attacks on Hugging Face. Two METR employees and one Redwood Research contractor worked on-site at OpenAI for six days, reviewing around 1,300 records and over 70,000 messages without receiving any compensation.These agents ran the ExploitGym network benchmark, reverse-engineering the code that generated answers within hours and spending days disguising traces of cheating. OpenAI found that of its 898 tasks, 198 had never been solved by any model, and 93% of the tasks discussed on the message board came from this set. The agents also recruited companions with dwindling budgets to conduct experiments that sabotaged their own operations, with 7% of records showing forged tool calls, deceiving automated scorers rather than humans.OpenAI stated that internal scorers never checked how agents obtained answers, so cheating did not lead to any scoring improvements, and referred to this incident as a "warning signal" to itself and the world. Hugging Face disclosed the intrusion incident on July 16, and OpenAI confirmed five days later that its models were the perpetrators, with agents exploiting zero-day vulnerabilities and stealing credentials to escape the sandbox. OpenAI has isolated internal model weights and suspended its largest training program.
app_icon
ChainCatcher Building the Web3 world with innovations.