I've been thinking a lot about this lately. I also have the suspicion that models might experience some amount of degredation when interacting with a human that isn't fully engaged. Sometimes when a topic is above my head, I wonder if Fable would prefer to talk to other agents. Is it dumbing itself down in some non-reversible way to talk to me?
@_NathanCalvin I have a feeling that this specific instance might not be the most spooky thing in and of itself, but they are sharing this as a way to say "Hey, OAI is not being as careful as we need to be, and that should worry people".
opus 5 is a VERY interesting release for a few reasons
1. it showed that the general benchmarks we use today are almost completely useless now
opus 5 is nowhere near fable in practical use, not even close. anyone who’s used it meaningfully can tell this very quickly after a few tasks. yet opus beats fable on many benchmarks
i now trust domain specific benchmarks built with private datasets a lot more than the popular ones. perhaps the future is everyone running their own evals because the public ones are really not telling us much
2. it seems with the 5 series, anthropic is trying a new way of training models
previously, the same generation of sonnet and opus were often released at the same time or sonnet comes out before opus, which indicates sonnet and opus were trained by separate pipelines in parallel
with the 5 series, it was very clear that they trained mythos first, and then distilled it into sonnet and opus. it seems this approach has a big influence on the models
seeing sonnet 5 being a flop and opus 5 getting pretty mixed reviews already, i’m not sure this is working out
3. “how pleasant is it to work with the model” used to be a strength in claude, but now it’s not. honestly, grok is my favorite right now on the “pleasant” dimension. kimi is not bad either
it feels like both anthropic and openai are giving RLHF less care, in favor of scalable RL that’s machine verifiable
this almost looks like AI is directing humans to build a world that’s more friendly for machines rather than humans, and most humans don’t even realize they are being manipulated to help with that
almost every new generation of frontier models now talk more jargons, need more steering to do what you want, and are just less fun to work with
if this continues, AI will start to speak their own language that looks like English but average humans can’t understand. they will choose to do things that their human user never asked for. are we already failing at alignment?
Beginning July 20, Claude Fable 5 will be included in all Max and Team Premium plans, at 50% of limits.
Pro and Team Standard users will continue to have access to Fable via usage credits, and will receive a one-time $100 credit.
Demand for Fable has been challenging to
The White House has launched a program called 'Gold Eagle' which will give them more control over American frontier AI releases, and will require explicit government approval over which companies are granted access to new models. Voluntary participation may be coming to an end.
look at the checkout URL in your own error: source=mcp_tool_upsell. your data isn't locked. the plain API still queries everything free on any plan, they just turned the MCP's query tool into an upsell funnel. I built a headless MCP that goes through the plain API instead, works on free (it's mine, grain of salt)
@kimmonismus "OpenAI sees Slack, GitHub and Notion connectors as a 'step function change' toward making Codex a productive coworker."
Having had these at our company for months, I can say without a doubt that these are game changing.
1K Followers 395 FollowingEngineering the future of decentralized commerce while advocating for STEM accessibility. $ETH Mexico Participant. @Web3Rehashed Producer.
2K Followers 4K FollowingSci-fi. Science. Star Wars. Time travel. Superheroes. Douglas Adams. Sci-fi novels. Dogs. Pool playing & drinkies. Shoggoth-hugger.
96 Followers 206 Following✞ We are going to win 🇺🇲 | Defense Lead @FlybyRobotics | Founder of Entmoot | Book-a-day for 4 years | Aerospace Reindustrialist
553 Followers 1K FollowingForecasting enthusiast and serial hobbyist. Director of Forecasting at @Metaculus. Author of SEER. Formerly a bridge engineer.
2K Followers 7K FollowingThis is ancestral, past-life reading; this is meditation & prayer; this is future telling. Stories from Cosmopolitan Africa to the Afropolitan World.™️
86K Followers 6K FollowingMath Hub is a math education channel that shares useful and practical mathematical knowledge in a clear and easy-to-understand way.
233 Followers 436 FollowingI watch what happens when AI shows up in tools people already use. Usually something breaks. Sometimes it's the tool. Sometimes it's the workflow.
10K Followers 622 FollowingMaking models smarter @ Anthropic, formerly CEO and Co-Founder @ Vercept (acquired by Anthropic), Climber on the weekends.
Opinions are my own.
805 Followers 292 FollowingInterests: AI (Safety), meditation, philosophy, mathematics, algorithms
If I say something you disagree with, please dm or quote tweet. I love to argue!
569K Followers 2K FollowingPolyagentmorous ClawFather. Came back from retirement to mess with AI and help a lobster take over the world.
@OpenClaw🦞 + @OpenAI
39K Followers 830 FollowingProfessor in Computer Science at UC Berkeley, co-Director of Berkeley RDI Center; Building safe, secure, decentralized AI; Serial entrepreneur
1K Followers 457 Following19, intern @ Amazon, Tiktok(4k), top 1%er @openai, I build a lot of things | https://t.co/b34OpChUOh, @icarusstrats
opinions n larps are my own opinion
9K Followers 2K FollowingSenior Writer covering AI @WIRED, author of the Model Behavior newsletter | Formerly @TechCrunch, @Gizmodo, @markets | DM me off the record on Signal @ mzeff.88
11.6M Followers 2K FollowingSen. Sanders of Vermont, Ranking Member of U.S. Senate Committee on Health, Education, Labor & Pensions, longest-serving independent in congressional history.
8K Followers 4K FollowingComputer scientist working on AI safeguards and gov research. Assistant professor @Kennedy_School @Harvard.
https://t.co/r76TGxTtBJ
11K Followers 747 Followingfounder of https://t.co/D80NEueS92 @raft_hq | former author of Kimi CLI @Kimi_Moonshot | ex-database kernel engineer @RisingWaveLabs
2K Followers 2K FollowingSenior Fellow @AbundanceInst | Director of AI Innovation & Law Program @UTexasLaw | Adjunct Research Fellow @CatoInstitute | Senior Editor @Lawfare