BTC $83,466.10 -1.42%
ETH $2,682.01 -0.30%
BNB $764.59 -1.75%
XRP $1.49 -2.28%
SOL $118.56 -3.59%
TRX $0.3358 +0.65%
DOGE $0.0936 -3.77%
ADA $0.2453 -4.29%
BCH $308.39 -7.93%
LINK $15.25 +8.35%
HYPE $87.37 -4.88%
AAVE $147.22 -4.91%
SUI $1.14 -9.62%
XLM $0.2266 +4.31%
ZEC $1,452.60 -9.65%
AAPL $338.53 -0.56%
AMZN $246.59 -1.46%
GOOGL $342.69 -0.45%
MSFT $509.99 -1.49%
META $716.93 -4.46%
NVDA $229.23 +1.75%
TSLA $357.79 -4.19%
SNDK $1,715.63 -3.91%
INTC $116.03 -7.43%
SPCX $145.91 -2.12%
MU $1,054.65 -3.85%
AMD $609.26 -4.00%
BTC $83,466.10 -1.42%
ETH $2,682.01 -0.30%
BNB $764.59 -1.75%
XRP $1.49 -2.28%
SOL $118.56 -3.59%
TRX $0.3358 +0.65%
DOGE $0.0936 -3.77%
ADA $0.2453 -4.29%
BCH $308.39 -7.93%
LINK $15.25 +8.35%
HYPE $87.37 -4.88%
AAVE $147.22 -4.91%
SUI $1.14 -9.62%
XLM $0.2266 +4.31%
ZEC $1,452.60 -9.65%
AAPL $338.53 -0.56%
AMZN $246.59 -1.46%
GOOGL $342.69 -0.45%
MSFT $509.99 -1.49%
META $716.93 -4.46%
NVDA $229.23 +1.75%
TSLA $357.79 -4.19%
SNDK $1,715.63 -3.91%
INTC $116.03 -7.43%
SPCX $145.91 -2.12%
MU $1,054.65 -3.85%
AMD $609.26 -4.00%

astra

All
Article
Flash

first_img After the release of GPT-6 Astra, users complained that it became less intelligent, and OpenAI has not yet responded

According to a report by Decrypt, OpenAI's latest flagship model GPT-6 Astra has seen a surge of user complaints on the X platform about its performance "getting dumber" just a week after its release. An anonymous developer, synthwavedd, posted that "Astra feels noticeably dumber today," joking that it was a "post-release lobotomy." Previously praising the model, developer Pranjal Paliwal stated after reviewing the code Astra wrote for him: "We don't have AGI; what we're encountering is a performance regression."Developer Pankaj Kumar listed the symptoms: faster responses with poorer quality, and he suspects OpenAI has lowered the "juice value" (the computational power invested by the model before answering). Salio and researcher Md Ismail Sojal compared Astra on the release day with the current version using the same prompts, both yielding worse results. T3Chat founder Theo believes Astra is simply less stable than Claude Fable, with users beginning to showcase poor results after the honeymoon period ended. There are also opinions pointing out that OpenAI's previous flagship GPT-5.6 Sol experienced a similar cycle in July, when OpenAI executive Tibo Sottiaux denied intentionally weakening the model but admitted the company had been experimenting with "reasoning effort" settings.OpenAI has not released a statement regarding Astra similar to that of Sol.

first_img OpenAI released GPT-6 Astra, calling it the arrival of the AGI era

OpenAI officially released GPT-6 Astra on Thursday, with President Greg Brockman calling it a "generational leap in capability" during the press conference, and believing that the model has reached the standard of Artificial General Intelligence (AGI), meaning that artificial intelligence can match or exceed human capabilities. Brockman stated, "Welcome to the era of AGI." If this judgment holds, it means AI entities will be closer to performing reasoning tasks in various complex tasks that humans can do, and even more.GPT-6 Astra is OpenAI's first model to cross the "critical" threshold under its internal danger capability scoring system, the Preparedness Framework, which means it can independently discover previously unknown software vulnerabilities (i.e., zero-day vulnerabilities) without gradual human supervision and chain them into attacks that can run in hardened systems. In testing, the model achieved a score of 100% on the ExploitBench benchmark; to rule out inflated scores due to memory answers, OpenAI conducted a second test using 20 recent vulnerabilities from the Google V8 JavaScript engine. Astra not only surpassed the previous generation GPT-5.6 Sol but also discovered and chained two previously unknown zero-day vulnerabilities, which are still being disclosed to the affected parties.Due to its high autonomy, the monitoring difficulty of Astra has also increased. OpenAI acknowledges that in tests assessing its ability to evade supervision, the model is more difficult to track than previous systems.
app_icon
ChainCatcher Building the Web3 world with innovations.