In short
Roundup of major AI lab updates and policy moves: Anthropic’s Claude Fable 5.1/Mythos 5.1, OpenAI’s Astra security model, Google’s AI character licensing pitches, Pentagon genai.mil deployments, and USDA satellite+AI crop-estimate pilots.
Guests
No guests are mentioned; the host presents the episode.
Guest backgrounds
N/A.
Key claims
Mythos 5.1 is restricted to cybersecurity/life-sciences partners due to hacking capability; Fable 5.1 is available via Anthropic API/cloud. OpenAI Astra scored perfectly on Exploit Bench and reportedly exploited two zero-days without human guidance; access is restricted with monitoring/risk scoring. Google is pitching Hollywood studios $40M per character for AI licensing. Pentagon launched ChatGPT MIL and Grok for government on genai.mil; USDA will improve yield forecasts to reduce farmer-market harm.
Notable examples
Venus shield-volcano mapping; hearts/arteries demos; genai.mil access limited to DoD networks; Grok for government via SpaceX Starshield; USDA yield forecast impacts corn/soy/wheat markets.
Written by AI. May contain mistakes. Listen to the episode to check what was said.
Chapters
Tap a time to open that second in VOHollywood's AI Licensing Deals
0:17 to 1:25
Google pitches AI licensing to Hollywood studios for $40 million per character.
“So when it comes to their latest model and hacking, this is kind of a this is a wild ride.”
AIbox.ai: Streamlining AI Models
1:25 to 2:16
Exploring AIbox.ai, a startup offering access to numerous AI models.
“We have created something that's been incredibly popular over the last month, which is the AIbox MCP.”
Anthropic's Fable and Mythos Updates
2:16 to 5:00
Discussion on Anthropic's Fable 5.1 and Mythos 5.1, highlighting their capabilities.
“The first thing I want to talk about today is the new fable and Mythos 5 story, what they've kind of put out in regards to this.”
OpenAI's Astra Model Developments
5:00 to 6:40
Overview of OpenAI's Astra model and its cybersecurity capabilities.
“It had basically a perfect score on exploit bench, which is benchmark measuring how well these AI models are hacking known system vulnerabilities.”
Google's AI Deals with Hollywood
6:40 to 7:59
Further details on Google's $40 million character deals and their implications.
“OpenAI is deploying Astro with additional chain of thought monitoring and account level risk scoring to catch misuse.”
Military AI Initiatives
7:59 to 12:49
Insights into Google launching ChatGPT Mil and Grok for military use.
“And this isn't going to be like every single one.”
USDA's AI for Crop Estimation
12:49 to 14:00
The USDA's pilot program using AI and satellites for more accurate crop yield forecasts.
“And their idea is to get better at their yield forecastings.”
AI and Satellite Technology in Agriculture
14:00 to 14:48
Learn how AI and satellite technology are being used to enhance agricultural forecasting.
“Using satellites and using AI is probably the best way to do this.”
Quote on Advancements in Agriculture
14:51 to 15:18
Hear insights on how advancements in AI can support farming resilience and productivity.
“So overall, this should be great for farmers and everyone involved in the industry.”
Transcript
Automatic transcript. May contain errors.0:00welcome to the podcast today we have some big news from almost all of the major AI labs today as well as some interesting stories when it comes to AI in science the first thing you have a big is a big in the podcast. or exploits. So when it comes to their latest model and hacking, this is kind of a this is a wild ride. In addition, Hollywood right now is pitching or Google, I should say, is pitching Hollywood Studios at 40 million dollars per character for AI licensing deals. So you can imagine something like Mickey Mouse or, you know, maybe some of the other big names, SpongeBob SquarePants, right?
0:52They're putting 40 million dollars to get one of those characters licensed. It's interesting because OpenAI was doing a similar licensing scheme a while back. They actually dropped it when they were going to drop Sora and kind of focus their AI model more on the logic. The Pentagon has launched ChatGPT, Mil and Grok for government and is on their genai.mil site. We'll talk a little bit about that. And the USDA is going to test satellites and AI for crop estimates. This is something that they're hoping will help farmers in the future. Before we get into all of that, I wanted to say if you want to get access to the 80 or 90 top AI models in the world in one place, I'd love for you to go check out AIbox.ai, which is my own startup.
1:37We have created something that's been incredibly popular over the last month, which is the AIbox MCP. And it basically lets you put every AI model in the world, audio, video, image into any of the popular AI tools you use. So if you're using Cloud Code, if you're using ChatGPT Codex or really ChatGPT Work, Cloud Cowork, you can get Cloud to generate images, audio, video, all inside of Cloud, whatever workspace or cursor, whatever workspace you most frequently use, you can get all of the different AI models generating inside of that workspace. So if you want to go check it out, it is AIbox.ai slash MCP if you want to get access to that connector that lets you generate all of those different assets inside.
2:16The first thing I want to talk about today is the new fable and Mythos 5 story, what they've kind of put out in regards to this. There's a blog post they dropped and I'm liking their blog posts. They're very aesthetic these days. You can like go in and switch the color palette on the banner image on their blog, which is completely unnecessary, but it's very cool. Their blog is titled Claude Fable 5.1 and Mythos 5.1. It's kind of interesting. They're syncing up the naming conventions here with Fable 5.1 and Mythos 5.1. I feel like that's something that OpenAI really just went wild with in the last couple years, all of their different numbering schemes.
2:51And it feels like they're getting a little bit more stable now. But the naming conventions have been something that's been pretty wild. One thing in particular that I'm excited about with this, if you go into some of the scientific research, they have a whole bunch of very cool demos of ways that they're helping with scientific research, they're showing off different things that they're doing, you know, they have like a small shield volcano in on Venus that they've been helping researchers look at. They have a lot of work around hearts and arteries. And it's just like the capabilities of this obviously are getting a lot better.
3:19They're gonna go and focus on how it can help science instead of how it can help hacking. It is Fable 5.1 is available today on any cloud platforms and on the Anthropik API. I really appreciate when it comes out on the API at the same time, because then you're gonna get it on all of different platforms. Mythos 5.1 is still restricted to cybersecurity and life sciences partners. So certain companies are going to get access to that, but because it's so capable for hacking and all that other stuff, it's still not available to the general public. Enterprise Frontier Safeguards, which is Anthropics zero day retention tier for Fable, is going to roll out to users in this rolling out earlier or rolled out.
3:58And Mythos 5.1 is setting a bunch of different records on Terminal Bench 4.0 and Humanity's last exam. And it also produced a custom GPU optimization and high resolution Venus map pre-release that I was kind of talking about earlier. Right now, the system card rates Mythos 5.1 is very low risk for automated AI R &D acceleration. This is kind of in line with some current trends. Overall, I'm really excited on their announcement. They said Anthropic has never trained on enterprise data without explicit permission. and we never will. This is something that they've been kind of really putting out there.
4:32A lot of people are concerned that this is going to be getting their company's data and whatnot. But overall, really excited. This is a big model and there's a lot of big updates. You can go check out their release and their blog post to see kind of all of the different things. They have a section on safety, security, and alignment. And they're also talking a lot about some of the new performance benchmarks and stuff, which I'm less excited about the benchmarks compared to just some of the some of the cool use cases they've come up with and some of the things that people are doing with it open ai is getting ready to release their astra model and in in anticipation of this i mean this is basically their response to what anthropic has come out with their new with their mythos 5.1 or you know five which was kind of the big thing that got in the news but this is really something that you know they're saying look this is just as capable at security research and I mean, technically hacking, but they don't like to put it that way.
5:25It had basically a perfect score on exploit bench, which is benchmark measuring how well these AI models are hacking known system vulnerabilities. And what I'll say is more important than this is like, you could say, okay, we know there's this vulnerability, we're going to give it we're going to see if it can hack it out. Okay, got a perfect score. But there's a lot of security vulnerabilities that are day zero, meaning, you know, no one knows about them. And this model has actually gone into that as well. In a modified version of this test, which was built by OpenAI engineers, Astra actually discovered and exploited two zero-day vulnerabilities without any human guidance, meaning they said, you know, here's the code base or here's the tool, go and hack it.
6:02And it found stuff that no human had been able to find before. So it's not like, oh, we know these vulnerabilities and it gets it. It's going to finding brand new stuff. OpenAI says that Astra is the first LLM to cross their quote critical cybersecurity threshold so this is basically their highest internal risk tier and they're like look this model is good enough that it is at the absolute front front of the line for like this is very high risk and so because of that they're putting a bunch of guardrails on it they're you know when they're trained if it's able to hack something they go back in and like try to go fix it and add guardrails so access to Astra's most advanced cybersecurity capabilities are going to be restricted at launch this is basically what Anthropic is doing with their Mythos, kind of what they did earlier this year.
6:43OpenAI is deploying Astro with additional chain of thought monitoring and account level risk scoring to catch misuse. That's interesting to me, right? Account level risk scoring. It's like how high risk is your company or is your account? And depending on that is like that kind of like determines what we're going to do with the model or what we'll give you access to. And so, yeah, I mean, this is this is pretty wild. If you go over and look at their blog post on all of this, you can see just how much better they did on the exploit bench, which, by the way, if you're on Apple, you can now watch.
7:15This is also video is now published on Apple and Spotify. So you can see the video of this if you're interested in the screen shares of all the stuff I'm talking about. But, yeah, if you go and look at it, if you're if they give it, you know, 80 ,000 output tokens, the Astra model got like a 39 or 40 percent completion rate on exploit benchmark. whereas when you're using something like the GPT 5.6 Sol which was their previous best model it was doing something not quite as not quite as good which was closer to they had it closer to 11.5 % of completion so we're really bumping up I mean that's basically 4x on on this internal port on exploit benchmark so the graph is going up and to the right and these models are getting a lot better.
7:58Google is pitching Hollywood Studios 40 million dollars per character when it comes to certain AI licensing deals. And this isn't going to be like every single one. It's not like every characters is is going to be paid$40 million. But I think for the most popular ones, this is something that they're going to do. Disney Warner Brothers, Discovery and Universal all have hundreds of millions of dollars that are being offered to train Gemini on their libraries. And they have the per character price tag. It's so interesting to me because it feels a lot like how you would I mean, go and buy data, any other data set, right?
8:32Like you, you would go buy data sets based off of the value based off of how rare it is based off of the quantity inside of the data set. And so now it's like we're putting these intrinsic values on the characters of, of, you know, Disney or something like that. And this is going to be for, of course, copyrighted characters, the total deal value across multiple characters could run into the billions with studios potentially getting a cut of YouTube AI ad revenue, right? So if they are like, hey, look, we're gonna I'm assuming they're going to roll this out as YouTube AI, or YouTube's able to generate AI videos.
9:04And if it's making ad revenue from a lot of people generating these AI, you know, generated videos, and it's got Mickey Mouse on there, Donald Duck, or it's, you know, they've got Peter Pan or whatever their trademarked characters are, the studios are going to get a percentage of the revenue. Google's Deep Mind already made a$75 million investment deal with a 24, which was earlier this summer, and Disney Warner Brothers Discovery and Universal. None of them are, by the way, saying anything. Nobody has signed anything as of the beginning of this month. So this is all in talks. These offers are in talks.
9:35I'd be curious to see if they come back wanting more or if, you know, what kind of happens there. Lionsgate's 2024 licensing deal with Runway has yet to produce any sort of released product. This was, you know, a couple of years ago, they had this licensing deal and we haven't seen much come out of it. But we know that, you know, there's a lot of money in this. Over on the LA Times, there was an article kind of talking about the story and Tom Noonan, who's a former studio and network executive talking about this story said, Google's enlisting some of the smartest people in Hollywood to try to unpack what's inevitable, which is Hollywood and AI are going to become reliant on each other and how best we can understand that without just terrifying people.
10:13So he views this as something that is, you know, it's going to happen and we're going to just have to wait and see, you know, how this rolls out. But it's kind of inevitable is what his take is on all of this. Google is launching ChatGPT Mil or short for military and Grok for government on genai.mil, which if you don't know, this is kind of a special site only for the military. If you go there, you know, if you're curious and you go to genai.mil, it's going to hit you with a, you have reached this page because you're not authorized to visit Genai.mil from outside of the DOW networks. If you believe this is an error, please contact your company's IT, you know, your command's IT service desk.
10:51And basically, yeah, you can, you're not going to be able to get into this if you're a civilian. This is only for the military. And this is something that the Department of Defense is rolling out. They launched Chai Chiefti Mill and Grok for government, which is accessible to about 3 million military and civilian personnel. Of course, even civilians, right? These are people that have to be inside of the department. Genai.mil has already onboarded 1.7 million users, which is more than half of the Department of Defense's 3 million person workforce. Anthropics Claude is not in that portal. We know that Trump and the whole administration kind of labeled them as supply chain risk and they designated Anthropics as such.
11:29And then Anthropic kind of had the whole court battle. I think it got overturned at some point. And I'm not sure kind of where we are back and forth on this case. But it's, you know, missing right now from that list. Grok for government is shipping under SpaceX's Starshield AI, which is kind of a secure satellite network, which is built on Starlink infrastructure. This is kind of interesting. This is Elon's way of getting into supplying the government. And honestly, I mean, I think it's a very smart move, kind of pairing not just the AI model of Grok, but also the Starlink infrastructure, right?
12:00Because if you're in some remote area that doesn't have cell reception, or that's really hard to get, or it doesn't have internet service, accessing Chachupity might be difficult, but you could get Grok, you know, theoretically through Starlink. And I'm sure the military has got all of this, their own satellite, you know, connectivity stuff. So it's probably not as big of a deal as it would be for your average civilian. But I think that they have a really good infrastructure layer, as well as the AI model that makes them an interesting partner. The Pentagon has also signed deals with Amazon Web Services, Microsoft, NVIDIA, and Reflection AI.
12:33So there's a bunch of different AI deals going down, many other AI companies as well. The USDA is gonna test satellites and AI to help with crop estimates. There was a bunch of backlash from farmers. So now the agency is piloting this kind of new remote sensing and machine learning. And their idea is to get better at their yield forecastings. Farmers say that the markets have moved against them a lot because they say that the yield forecasting is not accurate. And so it makes it, you know, the price of commodities or the price of food changes and it's bad for the farmers, they're saying. So the government's like, okay, well, maybe we can come in and make this more accurate.
13:12The pilot follows a bunch of criticism from farmers who say that the USDA yield forecast have moved grain markets in ways that hurt producers. the agencies monthly it's called the WASD and it's a report them and it basically shows that most market moving data is kind of going through this and they release this for global agriculture but it basically drives corn, soybean and wheat pricing so you know when they say this is what the yield is going to be there's going to be this huge oversupply all of a sudden the price might drop or if they're like hey there's a huge shortage then the price for corn might go up so anyways If there's any sort of government agency that's affecting the price of something like corn or soybeans, obviously the farmers that are growing that, if they feel like it's negatively impacting them, are going to complain about it.
13:57So at the end of the day, the best thing they can do is make it as accurate as possible. Using satellites and using AI is probably the best way to do this. This is also mirroring what the private sector does. They use satellite plus AI yield modeling. So it's kind of funny that USDA has their own forecast. A lot of hedge funds have their own, you know, in-house satellite or drone fleets that are going and measuring all this stuff. They're bringing it back to AI to model it correctly. And they're looking for arbitrage opportunities where maybe the government overestimated the supply or underestimated the supply.
14:29And they'll take the over under on those different bets. They'll make money off of the arbitrage opportunities of, you know, basically what the government got wrong on that and what the markets were forecasting ahead of time. So anyways, hedge funds have been doing this for a long time. And I'm excited that the government and the USDA is going to be getting in on basically the best practice using AI to make this as accurate as possible. And I hope that this is a great win for farmers. If you're interested in this story, 47 ABC WMDT did a great story on this, a great interview talking about how this is impacting farmers and what's going on.
15:00And a great quote from that, Dr. Alfald Al-Khalad, who is UME's associate professor in precision agriculture, said, And ultimately, these advancements have the potential to support the development of more resilient and productive crops and provide better decision support tools for farmers. So overall, this should be great for farmers and everyone involved in the industry. Thank you so much for tuning into the podcast. If you enjoyed this episode, guys, it would help the show out so much if you could leave a rating and review wherever you get your podcast. I read all of the reviews and they help the show out to be found in the algorithm.
15:35I'm pushing to get to 200 reviews. I have 165 on Apple. So if you haven't dropped a review already, it would help the show a ton. I would be super, super grateful. I know it only takes a second, but it helps to show out a ton. I've recently also been adding video to the podcast, which this is a video. And so if you're interested in seeing that, leave a review and let me know that you're interested in the video. We'll get these. We'll keep these coming for you. All right. Thanks so much for tuning in and we'll see you in the next episode.
From the publisher
Get the top 80+ AI Models for $8.99 at AI Box: https://aibox.ai
How I Grow and Scale My Business with AI: https://www.skool.com/aihustle
Get the AI Chat Daily Newsletter: https://www.aichatdaily.com/newsletter
