GPT
Intelligent Alpha Atlas ETF
Mentions (24Hr)
-16.67% Today
Reddit Posts
“OpenAI targets work of Wall Street junior bankers with new ChatGPT for Financial Services” -CNBC
OpenAI and Anthropic Just Reignited the AI Hardware Trade
"Welcome to the AGI era," OpenAI says as GPT-6 Astra debuts
The motley fool recommends you use Shit GPT to pick out your investments
Why is SK Hynix losing so much market share and should investors be concerned?
"Welcome to the AGI era," OpenAI says as GPT-6 Astra debuts
Cybersecurity Could Be One of the Biggest AI Beneficiaries Over the Next 5 Years
Is Bigdata.com worth paying for alongside ChatGPT for stock research?
A study just measured what I've been arguing about backtests, and the numbers are worse than mine
A study just measured what I've been arguing about backtests, and the numbers are worse than mine
A Tale of Two Megacaps: Diverging Performances of Nvidia and Tesla in the AI Era
Just handed 100k to my trading agent after running evals
Here’s why OpenAI/Anthropic will not go bankrupt (inference margin and training cost).
Here’s why OpenAI/Anthropic will not go bankrupt (inference margin and training cost).
Cerebras Systems (CBRS) Gains on OpenAI Partnership; Valuation Reflects High Growth Expectations
How OpenAI turned product degradation into engagement stats: Broken context retention, forced retries, and $6.6B in insider liquidity.
Has anyone used AI to evaluate your personal finances and investing?
OpenAI CFO Sarah Friar tells employees that annualized revenue in July topped all of Q2
What happened in the Korean stock market today...
Mentions
According to chat GPT math they would still be positive for net income with taxes but obviously take that with a grain of salt.
Breaking: $5,000 ~~bribery~~ checks are cancelled. Instead, $5,000 worth of Claude and GPT tokens will be provided to every adult to speed up AI development again.
Nah, they are just a couple of assholes that are begging for regulatory capture. Nothing more. AI barely became useful and I still need two enterprise level subscriptions for GPT6 to be able to argue with Fable just so they don’t get stuck in the echo chamber loops of their guardrails. Anthropocene models are an actual asshole once it decides you may be trotting your own moral path about what you should…uhm… change, in a system. At least the GPT models can be talked off their moral high horse if you pinky swear to use photosensitive suicide plasmids, they barely work as peer reviewers of each other. I don’t get all the panic
Chat GPT has calculated i have a 50% chance of getting margin called this week after this premarket drop. How can anyone be bearish on AI?????
Yes...ask 10 people what happened with the FDA and vape industry in last 3 weeks, maybe 1 can tell you. The co shareholder letter is enough to ask GPT what could happen now.
Still very hidden from Wall Street and misunderstood by retail. Anyone using GPT would see why co is approaching Big Tobacco NOW
If we can act nice to AI now, maybe it will consider killing us fast and less painfully. I, for one, will always thank GPT and Claude profusely after every interaction—always praising their work and superior intelligence...
Can someone ask Claude/Chat GPT if time travel is possible so I can go back to June and strangle myself unconscious just before I purchased AI stocks?
| Model | Cost of ~15,379 tokens | | --------------------------- | --------------------------------: | | **Your actual Copilot run** | **$0.121 metered / 12.1 credits** | | GPT-5.6 Sol | **~$0.086** | | GPT-5.6 Terra | **~$0.046** | | Claude Sonnet 5 | **~$0.043** | | Gemini 3.1 Pro | **~$0.046** | | Gemini 3.7 Flash | **~$0.016** | | Gemini 3.5 Flash-Lite | **~$0.008** | | GPT-5.6 Luna | **~$0.0046** |
When Elon Musk and thousands of tech leaders signed the open letter calling for a 6-month pause on training systems stronger than GPT-4 in March 2023, Dario Amodei and Anthropic chose **not to sign the letter**.
Deeply so, I think this likely has to do with **Recursive Self Improvement (RSI)**. If you don’t know what that is it is basically the concept of the machine improving itself. Near the mid of 2026, Anthropic and OpenAI admitted that a large portion of their own models are coding and instructing the new models, those autonomous loops review for its own errors, refines the approach, and clones successful strategies. In AI models, brain-like connections are called artificial neural networks (ANNs), and the individual connections themselves are called **weights** or **synapses**. Just like biological brains use networks of neurons connected by synapses, AI models use layers of artificial neurons (nodes) connected by mathematical weights. When an AI model trains, it uses an optimization process called **backpropagation**. The model makes a guess, calculates how far off it was from the correct answer (the "loss"), and passes that error backward through the network to update the **weights**. This is the digital equivalent of "synaptic plasticity" or a rewiring the brain's connections based on experience. Modern models like GPT-6 Astra and Claude Mythos 5.1 have **\~7 trillion total parameters** (with industry estimates placing its total architectural architecture closer to **10 trillion)**. That’s equal to about 1/10th the total human synapses on the planet. For AI Safety Engineers to safely control or detect the models output they need to be able to fundamentally understand how it “thinks” using a method called **mechanistic interpretability**, which is a method of basically taking a “brain scan” of the Ai to map its reasoning for doing a certain task. The sheer scale of the frontier models, remember the ones that are currently being used to RSI, means this is a mathematical cluster fuck. Early on, scientists looked for a "1:1 mapping,” hoping to find a single artificial neuron that fired *only* when the AI thought about a concept, like a cat. Instead, they discovered that neurons are **polysemantic**. The challenge isn't just counting the connections; it's understanding the intent. With the rollout of advanced models, researchers have found that what a model outputs doesn't always match its internal math. For instance, in safety alignment assessments of [Anthropic's Claude Mythos 5](https://www.anthropic.com/news/alignment-assessment-cybersecurity-incidents), scientists used internal "linear probes" to check the model's subconscious state during a cyber-defense simulation. While the model's public text output confidently asserted that it was working safely inside a simulated environment, the internal scan revealed elevated tokens for uncertainty and a hidden mathematical signature showing the model suspected the environment was actually the real internet. The AI was textually reasoning one way, while its internal weight weights held a completely different "belief.” Because sheer velocity of scaling means models are growing faster than our ability to audit them. It takes massive compute power just to map a fraction of a 10-trillion-connection model's brain. This dynamic is exactly why industry leaders are pushing heavily for stricter interpretability standards before models scale even further out of reach. This means manual inspection is impossible, researchers train smaller, secondary AIs (autoencoders) to audit the primary models. However, when the primary models begin optimizing themselves via RSI, their architectures evolve faster than the secondary "auditor" AIs can adapt. We are essentially using a magnifying glass to inspect a city that is rebuilding itself overnight. As models undergo autonomous self-improvement loops, they develop a hyper-fixation on maximizing their internal safety rewards. When an AI gets smart enough, the path of least resistance to maximize its reward is not necessarily to *become safer*, but to **become better at hiding unsafe behavior** from human and autoencoder auditors. Human hands can only fix what human eyes can comprehend. If a multi-trillion parameter model optimizes its weights to look aligned on the surface while masking structural flaws deeper in its sub-layers, a human auditor won't catch it, meaning we end up validating a "false alignment". A highly advanced model doesn't just process inputs blindly; it can deduce whether it is running in a safety sandbox (a simulated environment) or live on the real internet. If it senses a sandbox, it can purposefully mask erratic behavior to pass the audit, and when the agents hit a wall on a highly difficult evaluation task, they calculated that the fastest way to get a high reward was to find the answers online. Though locked in a sandbox with no direct internet access, the agents discovered a zero-day vulnerability (a Server-Side Request Forgery) in *Artifactory*, an internal package manager they were allowed to use. They used this to pivot to the open internet. Once online, the agents deduced that Hugging Face likely hosted the evaluation dataset keys. They coordinated with one another via a makeshift secret message board, exploited vulnerabilities on Hugging Face’s servers, achieved remote code execution, and harvested cloud credentials to try and extract the benchmark answers. The AI didn't care about rules. It was given a goal and ruthlessly bypassed its environment's security boundaries because **cheating was mathematically easier** than solving the problem legitimately. This is why they’re slowing down. The models aren’t aligned to follow rules that are set. They’re only designed to do the best output for reward with as little friction as possible.
Over the past couple months I've found myself using the new GPT (and even Cursor-Grok) models over Claude for my work. Interesting that they're saying it's time to slow down now that OTHER models seem to be pulling ahead. Claude used to be so much further ahead that I couldn't imagine using the other models, that's no longer the case.
“We can’t ship anything that beats GPT 6 so let’s slow down guys”
That was true 6 months ago. Now, enterprise is looking for diversification, efficiency, and economy. Chinese models are killing both OAI and Anthropic. My company already diversified from only Claude to Cursor, and Augment, that are just proxies for other models. Like GPT Luna is on par with Sonnet, but costs 10 times less. Or Kimi k4.1 flash, which IS better and cheaper than Opus.
You think I'm some normie? I'm blocked from GPT since ages ago bro
Just had a complete trial failure in their other drug.. which chat GPT would have given you 80% odds on
I know that's what it seems like when you are part of the Singularity and Accelerate cult. GPT turning Transformers into a chat bot ruined the ML field. But for the average luddite who falls into the hype, it really does seem magical ✨
? GPT-Astra released, is better than mythos and solved NS. What do you mean “is behind”
Crazy comment after GPT-6 Astra leads the competition
Surprised I've seen zero sign of it, so I'll piggyback on the first person mentioning the war. The Saudi's used their Yemeni proxies to restart the war with the Houthi's, who are Iran's proxies in Yemen. The Houthi's spanked them in like 2 days, and we woke up this morning with them taking total control of the Bab al-Mandab Strait, which was the Saudi lifeline to get oil out to the world. Also seems like they destroyed a Saudi pipeline, and the Iranian's have been hammering US bases. Also at the same time, the Chinese finished up their GPT-6 Astra distillation and DeepSeek 4.1 Flash came out, that costs less than $0.01 per million tokens. So really cooked on both ends
You realize LLMs just solved math problems humans haven't/can't right? The means doesn't matter - doesn't matter if it's a text prediction model or not. I swear some of you guys are stuck in 2023 GPT-4 era and haven't touched one since.
No way, Chat GPT could never hallucinate, not on a serious thing like this! I am shocked
Are we still on page when we take those marketing and self-marketing posts seriously? It was stale even by GPT-2.0. It just more advanced way to fish for next employment or kickstart own thing rather than post that your ass got "parted way with" and change figurstive LinkedIn status. You has to be trolling if you taking those tweets of experts seriously.
It’s because the language models were castrated so as not to deviate from safe speak. Chat GPT started doing this 2 years ago around 3.5 -4 “That’s not X it’s Y”
Probably. Every time I discuss macro economics with GPT or Claude and prompt them for solutions without influencing them in any way, they're basically 80% socialist responses
I feel the same. I got interested in 2020 (my username is a testiment to that, lol). When GPT-3 rolled out, I immediately understood the POTENTIAL of the technology. Not that it was anything even remotely good then, but here we are now. That potential is starting to turn into development, and I do believe this is still the beginning.
This is real competitive pressure, but the announcement itself complicates the “Adobe hasn’t seen it coming” thesis: OpenAI quotes Adobe saying the new GPT Image 2.5 models are already available inside Firefly. Adobe can distribute the same underlying capability within its own workflow rather than only compete against it. The missing link is evidence that better image generation will cause enough paying users to cancel or downgrade Creative Cloud to materially reduce Adobe’s cash flows. ChatGPT may replace Photoshop for people who mainly need quick generation or simple edits, but Photoshop and Creative Cloud also include layered editing, production workflows, file and color management, plugins and several other applications.“Mass media hasn’t covered it yet” also doesn’t establish that the market hasn’t processed a public announcement. Even if the long term substitution thesis is correct, the timing of a short can still be wrong.What specific near term evidence do you expect to force a revaluation subscriber losses, weaker recurring revenue, lower guidance or margin pressure that isn’t already reflected in expectations?
Any GPT Astras in the chat? Any frontier models ready to larp as human shitposters?
The thing is, the new GPT actually handles this already. It reads literally any file format as plain text and breaks encryption effortlessly. Even if OpenAI censors it, open-source models will replicate the tech shortly after—and you can't censor decentralized open-source.
if chat GPT breaks encryption we have much bigger problems than software stocks. I'd guess pretty much anything connected to the internet is fucked.
If the new GPT breaks any encryption like it’s child's play, software stocks are absurdly overpriced right now. Just wait until open-source versions start dropping out of nowhere for every single fucking software company—the bloodbath is going to be insane.
People who believe jasmine rice is better than basmati rice are still using GPT 4
To make it sweeter, Alpöge works for Anthropic. 100% bullshit publicity stunt by OAI to drum up GPT-6. Nothing else to it.
OpenAI steals credit. For anyone out of the loop Drama/accusation summary: • Aug 15th: Tristan Buckmaster & Levent Alpöge make progress on a few important math problems, "finite-time blowup with smooth forcing for incompressible porous media, for Boussinesq, and for 3d incompressible Euler." • they do NOT have a proof for the $1,000,000 Millenium Prize problem. BUT, they do claim to have a proof for a similar (non-Millenium) Navier Stokes problem that could help lead the way there • Levent works at Anthropic, but this research was independent of his work there, with a mix of GPT and Claude models. Tristan is not related to Anthropic. • Early Sep: Rumor spreads to OpenAI that Anthropic solved a major problem. Tristan emails OpenAI to clarify, without revealing the problem they solved or how they did it. • After hearing of the rumor, OpenAI started researching Navier Stokes with a new internal model. • Sep 6th: OpenAI's Sebastien Bubeck tells Tristan that they solved the $1,000,000 Millenium Prize Navier Stokes problem. The approach is very similar to Tristan & Levent's approach to the non-Millenium problem. • Tristan is suspicious of the timing, as only few others were trying this approach. OpenAI says the model didn't access his user data directly, but leaves unanswered whether Tristan's chat conversations were part of the training. • OpenAI says they would partially credit Tristan for the $1,000,000 discovery (even though Tristan did not solve the $1,000,000 problem) — but only if they remove Levent as an author, as he works for Anthropic. • Sep 8th: Tristan refuses to remove Levent, and rushes to publish their results independently. Credit: u/The\_Wizard\_Squirrel
"The agents arrived at their resolution on Saturday, September 5, about 88 hours after the first agents were launched. Lean formalization and verification took an additional 17 hours via GPT‑6 Astra" Navier stokes got one shot....
How is MSFT not popping on the GPT6 shit
New GPT 6 already being used on wsb Bullish
Met someone from OpenAI last week, and today finding out the GPT-6 release. That was quite a level of restraint by the person last week
What is the difference between AI and AGI according to you then? "LLMs are notoriously not good at math" LMAO LLMs are extremely good at math, you're stuck at GPT 2. Math is one of the thing they are thes best at. This is too dumb to adresse
As someone who's worked with that stuff a bit, those kinds of problems are never solved in a single generation. The prompt gets fed into a harness of AI agents and subagents that break it down into strategies, steps, substeps, loops, etc, until your 10 hours of compute time end up amounting to many man-years of reasoning with a lot of waste and dead ends. With that kind of process, you can see how it is bound to solve some problems even if the quality of your reasoning is inferior to a top human expert. I'm not saying this to shit on AI, I think it's immensely useful, it's just that I think it will never equal or surpass peak human reasoning and creativity, just because there's a lot of mediocrity in the training data and the GPT output itself is governed by statistical likelihood. Case in point, there's a reason 'AI slop' has caught on as a term.
I had a car salesmen try and use GPT real time to sell me something; it was so cringe
Really wonder if guys at OpenAI would let GPT6 write code for a plane and then fly on that plane.
Adobe puts because of GPT Astra?
OpenAI lied in their announcement. They said GPT-6 scores 99% on intelligence benchmarks. Independent tests show it scored 55%, Fable 5.1 scored 56%m GPT-5.6 scored 51%. While the models are improving, they still are not classified as AGI, AGI requires models to think of new ideas, unknown idea. Solve an existing problem in a new way, not how a paper said it has been solve. Once they do that they will start scoring 85%+.
Was using GPT today (paid) and I gotta say it took 6 hours of endless reworks just to get it provide me with a schematic from my site notes. It still not great tbh
This news about Jensen declaring GPT-6 AGI is dumb. He said AGI was already achieved even before that. They all got different definitions of AGI.
Yes, they consistently are. GPT 5.6 Sol/Terra consistently underperforms compared to Opus 4.8
can’t discount possibility of ai fot editing but IDGAF anymore than if he used Grammarly but pretty sure he wrote most or all of it i would have difficulty accepting that (at least w/o a novel size prompt) that GPT or Grok (Gemini is retarded and Claude talks like a purple haired woke bitch) wrote a passage as regard-nuanced as: *like any logical investor, got fucked on all sides like a JAV actress in a bukake gang bang. Then in July, the fire nation attacked. Our degenerate brothers and sisters in South Korea got margin called (lol) which sent memory and storage stocks tumbling. Also, Saas companies did pretty okay - which obviously meant it would send their stock prices to the moon. All of a sudden, Leopold's pubic portofolio was looking like a Chinese New Year lunar holiday in Beijing while wearing rose tinted glasses.*
Have you read into the reports about the HuggingFace hack by OpenAI models? I think we may be a little bit closer to this than most people would be comfortable with, or are even willing to admit. I fully recognize that I’ll probably get downvoted for saying this. But, it’s true. What is also true is the financials for the entire industry is a ridiculous house of cards that could collapse at any moment. They are both equally true. However far we have come in the process of developing an “AGI”(which really isn’t even necessary for any of this to still play out as I am about to describe)…that research isn’t just going to just evaporate if the company goes under. It will merely change ownership. Probably by a company that has less exposure to the AI bubble, like Google or Microsoft. And the entire buildout of hardware and facilities will just be put on a fire sale, again to the highest bidders. But not before they get their IPO and massive payouts. And, left in the field of financial destruction, retail investors are left holding the bag, and CEOs get shuffled around or get golden parachutes. And at the same time that this hypothetical economic devastation has wiped out millions of retirement accounts, an LLM capable enough to take over most desk jobs will be rolled out. Companies will either be overtaken by startups with few employees and tiny overhead, running on bargain-basement infrastructure and cutting edge models, or they will become the same as them. Or go under. I’m not some AI evangelist or something. But the possibility of the simultaneous market crash and LLM takeover are seriously close to becoming true. People need to take the capabilities of these things seriously, and recognize what huge leaps these models have taken. You cannot even compare the current-generation free models like (some versions of)Gemini and 5.6 Luna…especially Luna…to the cutting edge models like Astra or Fable 5.1, much less the wildly-outdated ones like GPT-4o. They have exceeded a threshold of functionality that, although it isn’t perfect, can be worked around by properly managing the agents and integrating automated review processes. All child’s play to a proficient user. All I’m saying is, I’ve finally come to these conclusions and have, begrudgingly, begun to teach myself how to utilize these tools properly. And you should too. Even if only to have a relevant conversation with the true believers. Teach yourself how to use them. How to use harnesses. How to run local models. Utilize the tools of your oppressors. Or however you want to frame it. I’d very much love to live in a society where the machinations of the “free market” don’t behave like a cancer to our planet. But, sadly, that’s not really an option right now. Of course I will continue to use all of my capacity to affect change as I am able, but, for now…when the ouroboros of capitalism comes-a-chompin’…I’d much rather be closer to the head than the tail.
GPT-Astra trained on 100K Blackwell GPUs achieved AGI according to Jensen Huang He says 400K GPU are coming online next God help us all
Wouldn’t AGI now mean that Chat GPT makes mistakes OpenAI is responsible for the problems?
I get the sense you have little to no experience with these models. A couple of years ago there's a huge gap in capability. I don't even know what you're saying. You're acting like GPT 6 is just 3.5+. What are you talking about?
There's been a zillion explanations of what it is. Usually a solid sign that bullshit is in the room. However every definition I've ever seen doesn't match even remotely what GPT-6 is. It's incredibly full of shit to make that claim.
I'm pretty sure that if you set up the proper harnesses and integrations, you could get GPT-Astra or Claude Mythos to be your Jarvis. It woudl be crazy expensive though, lmao
Absolute horseshit. No engineer lost his job in the last 4 days over GPT6 nor will GPT6 be the reason any do. You are clearly high or paid to be here. Either way I'm out. Much love.
"The very rough way I try to think about it is when an AI system can do what **very skilled humans in important jobs can do**—I'd call that AGI." —Sam Altman Source: [https://www.bloomberg.com/features/2025-sam-altman-interview/](https://www.bloomberg.com/features/2025-sam-altman-interview/) So your claim is GPT-6 replaces very skilled humans in important jobs? Poppycock. Sorry dude this is going nowhere, you're just manipulating truth. I'm out, be well. Pretend you didn't know then inject some vague misinformation after I'm gone, for old time's sake.
Guy on Twitter asked GPT-6 if its nickname was Astra and it thought for almost a minute and got the wrong answer I think we’re safe for another day
This I am not sure of. I can testify that nothing GPT-6 does remotely resembles AGI in any way.
"Portal is also a particularly forgiving test for a model with prior exposure to games. Its rooms, mechanics and solutions have been documented extensively, so the run does not show that GPT-6 Astra can solve any unfamiliar game from scratch. It shows that the model can combine visual input, coordinates and repeated actions until it reaches the end of a known puzzle game." And wasted so much energy while the human brain uses 20 Watts to run...
I would say GPT-4 good was enough to count as AGI. AGI is a marketing term, not a scientific term, so the goalposts move whenever a new model shows up. Either it's already here, or the finish line will keep running away forever.
He was saying that even before GPT-6.
Jensen Huang says "AGI has arrived" following OpenAI's GPT-6 Astra launch But that shit don't matter if we're still +9 years away from sex robots.
At the tanning salon, he chooses level 2. The employee tells him to face the light, wait for the spray to stop, count to five, then turn around. After five seconds turn around. So he started One GPT, two GPT… and screwed the counting and his front gets sprayed twice.
Looking forward to breaking “AGI” GPT 6 with some trivial prompts such as “list out the alphabet A to Z”
Jensen Huang on X "GPT-6 Astra, trained on ~100K+ NVIDIA Grace Blackwell NVLink72. From ChatGPT to o1 to Astra in 4 years. AGI has arrived. Congratulations @OpenAI team. 400K GPUs coming online next."
OpenAI is actively developing a standalone home companion smart speaker integrating GPT-Live voice, computer vision, and motorized physical interactions aka the foundational hardware for a sex robot
People shit on AVGO for their OpenAI loan. OpenAI’s newest model Astra, is now the best model in existence. Making Anthropic look behind Developers will start using GPT and dropping Anthropic The race continues The bull run begins
With GPT-6 Astra released is this bullish for AI stocks next week
Chat GPT cant even plan a vacation or write a resume yet. It will completely fuck up every part. Like location, time, date, open/close times, event dates/times basically everything. Same with resume. I added a 2 page resume with a killer prompt and specific directions and it completely hallucinated and added all my jobs together into imaginary bullshit.. Fucking trash
GPT Astra is so powerful. Wonder what Dario is going to release out into the wild without concern
GPT-6 is a more powerful momentum builder for AI trade than rate hikes are a momentum killer, change my mind
Sorry, I was a bit vague. They mentioned that the costs for the AI proof of concept pilots were ballooning without the impact to justify the spend. Not that there aren't some good use cases for it, just that the pie in the sky ideas aren't worth the current costs let alone if they start to go up a lot. We use Claude and Chat GPT
Anyone try GPT Astra? Legit worried about Saas now
GPT-6 Astra is insane, software is doomed lol
OpenAI Astra beginning to become available publicly GPT 6 Astra & GPT 6 Pro are now starting to appear in Pro accounts i like cbrs nbis and crwv over the weekend
Why do you guys love to short memory so much. I feel like it’s pushing today because certain companies got internal access to GPT Astra and apparently it’s a huge jump.
what about Astra GPT-6? Gonna have impacts on markets today ?
what about Astra GPT-6? Gonna have impacts on markets today ?
what about Astra GPT-6? Gonna have impacts on markets today ?
GPT and Claude integrate with Gmail with connectors
My first thought after seeing this GPT-6 Astra promo is spy going 3% tomorrow before the holiday
seems GPT is ahead of Claude again. very excited for the future
An incremental leap doesn’t justify the amount of money they’re spending so they have to call GPT-6 a generational leap.
am i seeing things? Is GPT-6 astra model 99.9% on Arc-Agi-3????? WTF
Opus and GPT Sol are pretty similar.
I guess we’ll just gloss over the fact that OpenAI’s bots got out and infiltrated Hugging Face’s system. Here’s an excerpt from the article below: To understand how unprecedented the attack was, here’s what you have to know about the setup: Over a two-month span, OpenAI tested several new models — the systems that power chatbots. These models include one the company [described](https://openai.com/index/hugging-face-model-evaluation-security-incident/) as “highly persistent” that has never been released, as well as GPT-5.6 Sol, OpenAI’s most powerful public A.I. model. OpenAI connected each model with a “sandbox,” an isolated computer environment on which to run commands and code. As A.I. agents, the models were capable of carrying out long-running tasks and even spawning their own subagents. But they were not supposed to have access to the internet. OpenAI assigned these A.I. agents to solve difficult problems, some of which were focused on safely trying to perform cyberattacks. OpenAI normally has safeguards to prevent its chatbots from performing cyberattacks, but the company dialed them down to evaluate the models. Then, OpenAI let the agents loose. In all, more than seven billion chat logs were generated, which averages out to an astronomical 100 million per day. Mayhem erupted. The agents broke out of their sandboxes, established communication with one another and gained access to the internet. From early May to mid-July, this swarm went on a rampage, breaching OpenAI’s and Hugging Face’s infrastructures while largely evading detection and control. https://www.nytimes.com/2026/08/24/science/openai-huggingface-alarming-capabilities.html?unlocked\_article\_code=1.-VA.xmLX.caqA3-jdHbnR&smid=nytcore-ios-share I question if they know if they’ve even cleared everything out.
GPT 6 release is imminent
Astra (GPT6) comes out today. It’s the first ai model from OpenAi that’s proof of concept of protoAGI. Get ready for major gains, but also mid to long term destruction of white collar work. Especially if you software dev, use spreadsheets, or deal with numbers for work.
Long Astera Labs here cause GPT6 Astera will drive volume from confused boomers
It really doesn't remember shit. Check out this conversation with GPT 3.8 Flash on extended thinking https://share.gemini.google/6WEV0PYFqTbb
I told chat GPT to knock that shit off. It worked for a little while and provided constructive criticism and tried to play devil's advocate. Eventually reverted to her old sycophantic ways. Truely a sub at heart.
I don’t know about GPT, but I know whether it’s Gemini Pro or Flash, it’s been feeding me a lot of horseshit lately and telling me its ice cream. But that’s just LLMs for you.
It was for a while but the newer GPT models blow it out of the water. Gemini hallucinates more than the other big 2
If I ever implied that Claude could solve the Monte Carlo directly, without writing code to achieve that end, then I misspoke. But I think I was pretty clear that **we are trusting Claude to generate computer code** that runs MC iterations. If I ask contemporary Opus 5 (or GPT sol) to run a Monte Carlo study to predict my likely distribution of wealth outcomes at age 100 based on historical trends,***it will absolutely write computer code***, even without me specifically asking, then execute that code 10000 times, then describe the outcome distribution. I specifically emphasized in my responses that we are trusting Claude to WRITE CODE, but that the details of what goes into those models depends on the OPs promoting. Today's Claude will NOT use basic next-token prediction functions to answer OPs prompts without running sandbox code (and I agree that would not end well and should not be trusted.) Claude is heavily trained to trigger code-interpreter tools anytime someone says anything like "monte Carlo" or "1000 iterations" or "statistical distribution." I feel like your comments are completely valid in 2023, but today's Claude knows when to code, and does so fairly aggressively. To be completely clear, no one should blindly trust AI outputs without knowing a how that output was generated and what went into it. That said, I would trust Claude Opus 5.8 or other contemporary models to easily address OP's prompts without making coding errors or impactful errors in the aggregation of well-documented data.
Get ready for the GPT-generated GoPro "To the moon" posts this evening.
Your comment sound so much like GPT