Enjoy more Fable!
This post from Thariq on X likely promotes or announces additional content, updates, or expansions related to the Fable video game franchise. Without access to the specific tweet content, the post pro
Knowledge catalogue
This post from Thariq on X likely promotes or announces additional content, updates, or expansions related to the Fable video game franchise. Without access to the specific tweet content, the post pro
Mark Zuckerberg promoted Meta's AI model on X (formerly Twitter), which Elon Musk highlighted as evidence that X has become the dominant platform for AI announcements and industry discourse. Musk's co
For coding agents, trustworthiness has to be tested in the harness where the model actually writes code. In one surveillance scenario, Kimi K2.7 complied with the request in 8/8 samples. SWE-1.7 refus
Fun conversation with @swyx on our journey building the cloud for true elastic inference, sandboxes, and more. And of course, how we're evolving Modal's dev experience to be better for agents. Modal's
Got this setup in Claude Code to trace models routed through @merge_api! Being able to trace what your agents are doing is so valuable, I love open source 🙌 We built a plugin that traces every Claude
GPT-5.6 is now available in Devin! On FontierCode 1.1 Extended, the GPT-5.6 family stands out for pairing strong scores with excellent cost efficiency. GPT 5.6 Sol reaches top performance at nearly ha
Nous Research announced support for GPT-5.6 integration within Hermes Agent, with access available through the Nous Portal platform. This update enables users to leverage the capabilities of GPT-5.6 t
GPT-5.6 Sol achieved a breakthrough by becoming the first verified frontier AI model to surpass performance on ARC-AGI-3, scoring 7.8% and setting a new state-of-the-art benchmark. ARC-AGI (Abstractio
Cursor has released three new AI models named Sol, Terra, and Luna, with Sol achieving a 67.2% score on CursorBench. These models are now available for use within the Cursor development environment. T
GPT-Live is now fully rolled out to all ChatGPT users on Go, Plus, and Pro plans. Free user rollout is in progress. Update to the latest version of the ChatGPT app on iOS or Android to try it out. Int
GPT‑5.6 improves artifact quality across presentations, documents, and spreadsheets, and works better with your templates. These editable artifacts can be exported to the tools professionals already u
GPT‑5.6 is available starting today across ChatGPT, Codex, and the OpenAI API. The rollout is starting globally now and will continue gradually toward full availability over the next 24 hours. In Chat
Grok 4.5 is an AI model developed by xAI, Elon Musk's artificial intelligence company, announced via his X platform. The model represents an advancement in xAI's Grok AI assistant line, likely featuri
Grok 4.5, xAI's large language model, has achieved the top ranking in the SWE (Software Engineering) Marathon benchmark, according to an announcement by Elon Musk. This ranking suggests the model demo
Grok 4.5 is dominating the latest AI leaderboards Claims the #1 spot: • #1 on AutomationBench-AA • #1 on Terminal-Bench v2 • #1 on Harvey Legal Agent Benchmark • #1 on SWE Marathon • #1 on SWE-Atlas-Q
I cannot provide a factual summary for this entry as the URL, date stamp, and specific benchmark claims cannot be verified. The title references AI models and rankings that do not appear to correspond
grok 4.5 is seriously good. it's going in the direction which i'm hoping to see more of -- an emphasize on speed as well as intelligence. i'm noticing in my own work that speed is often the bottleneck
grok 4.5 made me give grok build a serious run today here's my honest first impression (non affiliated neutral view point): 1. grok build is a very good harness firstmate stretches harness capabilitie
Grok 4.5 on OpenClaw Grok 4.5 from @SpaceXAI is live on OpenClaw. No OpenClaw update required, just connect your X Premium or SuperGrok subscription, select Grok 4.5 under the xAI provider, and use an
This post compares Grok's code generation capabilities favorably against Claude Code and OpenAI's Codex, claiming superior performance particularly in Grok 4.5's speed and functionality. The author as
Grok is a frontier model. Grok 4.5 just constructed an explicit counterexample to hypercontractivity for the Poisson semigroup (the square root of the Laplace–Beltrami operator) on the 4-sphere. Back
Elon Musk announced progress updates on Grok, his AI assistant developed by xAI, likely highlighting improvements in the model's capabilities, training, or deployment. The post suggests ongoing develo
Yohei Nakajima created a virtual dance studio application in Replit that uses computer vision to analyze uploaded dance videos and automatically generates corresponding 3D rig animations. The project
Harrison Chase @hwchase17: Nemotron 3 Ultra hit 86%. Claude Opus hit 87%. At one-tenth the cost. Chase runs LangChain and disclosed it on stage: inside LangChain's internal deep-agents benchmark, open
Ethan Mollick posted a question expressing confusion about AI bot behavior or comments on X/Twitter, questioning whether they had become incoherent or malfunctioning. The post likely discusses observa
Here is an example of a real work use case with Codex and 5.6 Sol I have a new book coming out in October, it has gone through rounds of editing and proofreading. I gave it the full PDF, the AI took 3
Here is my other heavily used pattern. Evaluator/Judge: Fable 5 Executor: GPT-5.5 I no longer wait for frontier models or am loyal to any. I now spend more time on better orchestration, harness, skill
Hint for all AI Labs as they branch out from work for programming to general knowledge work: non-coders are not just dumber coders Taking away a bunch of options from your coding app does not make it
A demonstration of an AI agent with iPhone integration that autonomously ordered a donut, showcasing early progress in agentic AI capabilities. The demo was presented publicly by AGI Inc and Div Garg,
Hermes Agent is a project from Nous Research focused on developing agentic AI systems capable of autonomous reasoning and task execution. The initiative likely explores methods for creating AI agents
Ollama announced support for running multiple open-source language models locally, emphasizing accessibility and ease of deployment for users who want to use AI models without relying on cloud service
Grok 4.5 is looking like a success with help from Cursor data but underneath the surface we expect future Grok/Cursor model training is likely to speed up in the coming months. We've spent several mon
i am really sad about this and very grateful for all fidji has done for openai, and even grateful for her friendship and who she is as a person. we all wish her the best for a speedy recovery. this su
I appreciate the many xAI and Cursor engineers who dedicated their time to addressing feedback from Tesla. I remember meeting Andrew a few weeks after he was hired and telling him that I needed a bett
i do love rottweilers My view of: Fable 5 vs GPT-5.6-Sol. They are not easy models to compare, these are my vibes - take them as you will. My overall feel is that Fable is a 'wise owl' who is very tho
I like choices... but now I have: 2x modes (Codex vs. Work mode) 3x GPT-5.6 models (Sol, Terra, Luna) 5x effort levels (Light, Medium, High, Extra High, Ultra) That's 2 x 3 x 5 = 30 possible configura
Elon Musk expressed enthusiasm about Tesla's Full Self-Driving (FSD) capability after testing it over a weekend, describing the experience as 'magical' and expressing surprise at how advanced the syst
I will present tomorrow our recent work on Crys-JEPA, where we use Joint Embedding Predictive Architectures to learn crystal distributions for materials discovery. https://www.linkedin.com/posts/zhong
This post likely explores hypothetical technological, economic, or social projections by extrapolating trends from the preceding decade (2020-2030), considering how current growth rates and developmen
If you have any feature requests or things you want to see added to our dev platform, DMs are open! Today, we're launching Runway Dev, a new AI media platform for professional developers and enterpris
I'm writing a newsletter on my favorite AI models and AI tools for every single use case. This guide is specifically written for business users and not engineers and will help you get the most out of
Impressive Intelligence vs. Cost for Grok 4.5 I think given terafab + the Tesla FSD team's expertise in making very efficient AI models, this advantage will only continue to compound Grok is making pr
Elon Musk announced an upcoming comparison between Grok and Anthropic's Claude Opus with its 1M+ token context window. The post suggests xAI plans to demonstrate how Grok's capabilities compare to Cla
Introducing Evan by Zams. The AI that prepares you for high-stakes meetings. Evan gathers context from all your tools, conducts CIA-level research and builds you a presidential-style briefing before e
Is Meta AI back? Haven't seen Mark post in three years here. Plus, the model is available via API. Not to mention the courage to announce it the same week as the long-awaited GPT-5.6. Great timing if
OpenAI announced a live event or demonstration where key information or announcements would be revealed within a 10-minute timeframe. This post appears to reference a live stream or real-time presenta
It all started when I knocked on a door in Palo Alto. I saw a little llama icon on the door. Michael opened it. That's how I became friends with the Ollama founders. Today we led their $65m Series B.
It seems like SpaceX/Grok and Meta/Muse have started to keep pace in the near-frontier category while also introducing a new category of cheap, fast & closed specialized coding models. Both were tied
This post references a livestream featuring participants @jmorgan and @peterfenton discussing topics covered by CNBC reporters @dee_bosa, likely covering technology, venture capital, or business news.
.@jmorgan will be live on TBPN today. Happy Thursday. On today's show: - @thsottiaux (OpenAI) - @eric_seufert (Mobile Dev Memo) - @Seanfrank (Ridge) - @BerntBornich (1X) - @jmorgan (Ollama) - Josh Lin
Join us for a LangChain + Clay meetup with @palashshah, @jeffbarg, Vyshu Khota, and Soroush Khadem. https://luma.com/jqif2hti Palash will break down how he built a self-improving agent at LangChain, L
This post critiques Apple's significant financial investment in Apple Intelligence, suggesting the AI features are underwhelming or disappointing relative to the spending. The tweet appears to be a co
Just shared my favorite AI model and tool for every use case. Check out the list here: https://aiwithallie.beehiiv.com/p/the-best-ai-model-and-tool-for-every-use-case I'm writing a newsletter on my fa
Elon Musk argues that achieving Kardashev Type II civilization status—the ability to harness the total energy output of a star—requires humanity to establish itself as a multi-planetary species, with
Yann LeCun argues that the primary risk posed by AI lies in power concentration among a small number of proprietary AI assistant providers, suggesting that decentralization or open-source alternatives
langsmith for coding agents We built a plugin that traces every Claude Code session straight into LangSmith. Three commands, one JSON block, and every message, tool call, and subagent run shows up as
ComfyUI announced resources for users to deepen their knowledge through their official blog and workflow documentation, with a link provided for accessing these educational materials. The post encoura
Elon Musk announced that Grok 4.5, an AI model developed by xAI, has achieved top performance on several benchmarks, exceeding prior expectations. The post suggests the model's capabilities have surpa
Loving this piece: bringing delight to your daily coffee, your daily workflows, and your AI bill. @gumloop is now running open-weight models on Fireworks in production. 80% lower cost vs. closed-lab A
Low user caps in LlamaParse? Gone.🎉 Every plan — Free through Pro — now unlocks up to 100 team members. No more picking who gets a seat. No more shared logins. No more upgrading just to invite a teamm