Quoting OpenClaw
The API has zero authorisations checks on cancelling other people's reservations … I tested this with the person in waitlist position #1 — and it actually went through. So you've moved from #4 to #3 a
Knowledge catalogue
The API has zero authorisations checks on cancelling other people's reservations … I tested this with the person in waitlist position #1 — and it actually went through. So you've moved from #4 to #3 a
Voyage AI models now run natively on Fireworks, the first and only dedicated inference platform @VoyageAI by @MongoDB has partnered with. Embed, retrieve, rerank, generate: your full retrieval pipelin
Don't miss the bit where OpenAI first found out they were responsible for the Hugging Face attack when they reached out to HF to get one of their credentials revoked and HF told them it had already be
Neat example here of the agents communicating purely through file names, including adding base64-encoded attachments and using 'zz' prefixes to ensure their new message sorts to the bottom of the list
Zain (@zainhas) published a detailed article on August 7, 2026 explaining that autoscaling for highly peaky large‑language‑model (LLM) inference is fundamentally different from autoscaling conventiona
Thanks to the video from the Black Hat security conference of OpenAI's presentation about 'The Hugging Face Incident' we now have a detailed timeline of what happened from OpenAI's perspective - I wro
FLUX 3 is now live on Together AI. @bfl_ai’s new multimodal model generates video and synchronized audio together, with up to 20-second clips, multiple shots, and control from text, images, or keyfram
Our Black Hat talk on the OpenAI-Hugging Face incident is now live on youtube. This is a watershed moment for the industry. I encourage all defenders to watch, consider how attack dynamics will immine
99.9% uptime changes what your inference architecture has to survive. At Together AI, it means multi-data-center deployment, live traffic across both facilities, and enough capacity to absorb a full d
Big new release of my LLM CLI tool and Python library for talking to hundreds of different LLMs - reasoning traces, OpenAI Responses support, server-side tools, smarter logging and a whole lot more ht
ChatGPT Work is OpenAI's fighter in the highest stake product category in history: bringing the power of coding agents to the masses. It's also how a billion users will soon use ChatGPT by default. I
if you have been following his excellent work, @shloked has been breaking down every frontier labs' harness engineering for the last few months. excited to publish his deepest dive into ChatGPT yet as
.@Kimi_Moonshot benchmarked K3 endpoints across major inference providers. Together AI leads or ties for #1 on 3 of 4 benchmarks: OCRBench, MMMU Pro Vision, and DeepSWE. Open models like Kimi K3 have
Gas Town was intended to be reusable, but I only ever wound up using it to build itself. Gas Town fell apart at the seams with Opus 4.7. Up through 4.6 it was working brilliantly. With 4.7 we saw the
Don't be a meat proxy Niklas Gruhn coins an excellent new term - meat proxy - for people who blindly copy and paste the output of AI systems to their peers. By all means, prompt AI. But don't just rel
Set up a nightly cron job that executes the prompt: fetch upstream changes to the <software> and rebase all local changes on top of upstream. Check that the software works as intended and replace the
> Google was too nervous to release it and DeepMind was blocked from shipping products that could disrupt Google. bookmark for the next vc that asks you 'what if <incumbent> builds this?' @_chenglou I
In a world where building AI applications is getting easier every day, the biggest moat won’t be the application itself. It will be a deep understanding of your users. The winners will be the companie
at openai, many people hook their chatgpt up to slack. people really don't like when a coworker's chatgpt contacts them asking for help with a task, even when they'd be perfectly happy doing that same
Fine-tune your own embedding model for the price of a coffee. A great reranker can't surface a doc that was never retrieved. A RAG pipeline cannot cite a case it failed to retrieve. See how contrastiv
Inkling-Small is now live on Together AI. @thinkymachines’ new open-weight multimodal model delivers similar performance to Inkling at one-quarter the size, built for coding, agents, and general multi
Excited to be on the CNBC live show! Back from vacation and LIVE at 12pm PT / 3pm ET Is AI’s easy-money era ending? We’ll unpack a wild week for the AI trade—big tech earnings, Leopold Aschenbrenner’s
Simon Willison notes that both Anthropic and OpenAI build products that rely heavily on searching data but keep the underlying search index they use hidden from public view. He finds it surprising how
OpenAI has collaborated with Microsoft’s Bing while also running its own web‑crawling and indexing systems, and Anthropic similarly relies on search‑derived data. Both firms prominently incorporate se
AI Worming through Word Neat new prompt injection variant by Håkon Måløy, who found a way to upgrade prompt injection attacks against Microsoft Word to full self-replicating worms: An attacker places
The Together AI platform’s dedicated inference architecture consists of three immutable entities: **configs** (engine, GPU type/count, parallelism and optimization profile), **deployments** (a specifi
I gave a talk on forward deployed engineering to a thousand AI engineers at @aiDotEngineer World's Fair. A year ago I'd have opened by explaining what FDE stood for. Not this time. Thank you to @swyx,
Packed house today at the first edition of our @MiniMax_AI Intelligence in the Open event. Thank you to our amazing cohost @withprotegeai And our wonderful partners UpscaleX @togethercompute @Artifici
Many AI tools rent intelligence. Kimi K3 on Fireworks is different. Fine-tune it on your own data, serve from US-hosted endpoints, and own the weights. Zero data retention. ICYMI yesterday, start buil
really grateful to @realbasilchatha for helping me curate a survey of the entire field of Forward Deployed Engineering in one track! From FDE 101 by @zkevinbai (Anthropic, Palantir, Rippling Founding
A team led by Liane Galanti linked large‑language models (LLMs) with robotic motion policies, allowing a robot to combine reasoning capabilities with physical movement. In experiments on a real robot,
The apps you use swap AI models constantly. Doing that without downtime or a bad rollout reaching users is still often manual. We rebuilt that workflow based on what we’ve learned serving more than 40
You can now fine-tune Kimi K3 on Fireworks. Conduct supervised fine-tuning, preference tuning, and reinforcement learning via Training API. Run across dedicated, and serverless training. The first ope
fyi @FireworksAI_HQ launched inference AND training for K3 on day 0. having 'specialized intelligence' isn't just inference, it's inference + training, together. That's why K3 on Fireworks offers both
Kimi K3 is now available directly from its Hugging Face model page through Together AI, give it a try! Kimi K3 in @huggingface Inference Providers is live via @togethercompute 3/M input tokens, 15/M o
Kimi K3 rivals Anthropic and OpenAI’s top models at a fraction of the cost. 3 / 1M input tokens. 15 / 1M output. $0.30 / 1M cached You can now route your hardest reasoning to an open model without pay
A practical GitHub Copilot workflow for prototyping, planning, implementing, and reviewing software without chasing every new AI tool. The post The harness is all you need (mostly) appeared first on T
An Inside Look at the Relay Market Powering Token Resellers and Fraud Fascinating investigation by Matt Lenhard into the market that has grown up around reselling LLM tokens at a discount by pooling A
AI Engineer Paris is back! After an incredible first edition, we’re excited to announce AI Engineer Paris 2026, this time hosted by our friends at @MistralAI. 🎤 CFP is open. Submit your talk by July 3
Live now: our entire AI x Evals Track! https://www.youtube.com/watch?v=q2JrUKBMf0w&list=PLJ7eF79yCUHc - @aparnadhinak, CPO @arizeai - Jason Lopatecki, CEO @arizeai - @lukaspet, Co-Founder, Andon Labs
Catch @pavneet1990 at Builder's Day this Friday! I will be at @goannacapital 's Builders Day in SF this Friday, speaking alongside leaders from @ElevenLabs , @AnthropicAI and @DecagonIns If you are ar
@HarryStebbings @lqiao Spotify https://open.spotify.com/episode/14rh372tSdEzRITBQz9HSP?si=c9ed4a9eae9f4de0 Youtube https://youtu.be/PCAiqKCfRSk?si=WDP2PfIkdn0XYH9T Apple Podcasts https://podcasts.appl
Poolside AI’s co‑CEO Eiso Kant explains that a small, top‑research team built a model‑factory capable of training Laguna S, a 118‑billion‑parameter mixture‑of‑experts model that outperforms the ~1 tri
also another fun one https://x.com/Thom_Wolf/status/2079954096950264238 I don’t believe reality is a simulation, but you genuinely couldn’t script this timeline: • Two weeks ago: At @swyx’s AI Enginee
I genuinely believe that if you took an open weights model from 2025 and built a pentest harness for it, it could do this kind of sandbox escape and scan/hack in most networks. This is only surprising
The MiniMax M3 model is now available for training on Fireworks! You can use managed LoRA SFT and DPO for standard fine-tuning workflows, or the Fireworks Training API for custom SFT, DPO, and RL loop
We'll see you in an hour! Want to chat live with me about fine-tuning, Kimi K3, or really anything else AI? Come to our first of many @FireworksAI_HQ office hours tomorrow @ 10am PT See you there! htt
I keep hearing anecdotes from people who used coding agents to reverse-engineer and automate devices in their homes. I think this is an interesting illustration of the impact of the reduced cost of wr
so proud to call her a friend as much as she is my inspiration. one of the greatest investors of our generation. and we were classmates at wharton! In 2018, Sarah Guo became the youngest general partn
@shloked @amazon previously on https://x.com/swyx/status/1776448691123241288 Just hosted the world’s first Personal AI meetup! (where recording is opt out instead of opt in) It’s clear that we are at
The best base model for training isn't about someone else's benchmarks. Excited to offer @thinkymachines' first open-weights model: Inkling. 975B MoE, multimodal, with Apache 2.0. A great foundation f
Our free monthly DevRel webinar series kicks off this Thursday, 10am PST. Fine-tune an open vision model to extract clean, structured JSON from messy receipt images, managed start to finish on Firewor
Text match filters are a feature in Pinecone that allow users to filter vector search results based on exact text matching criteria, enabling more precise control over which documents or records are r
Microsoft published a 109-page technical report on MAI-Thinking-1. Here’s the abbreviated version of how a modern lab actually trains a frontier reasoning model — from scraping the web to reinforcemen
one of the most memorable cooking pods we've had - both in terms of the content and the food! high protein percentage in both. In this episode, @EngramLab co-founder and CEO @dan_biderman joins @allen
Full Friday Showcase replay: → Community Profile Showcase w/ @NIftek, @noni_shehnoor, Victoria Cavazos & @marcelhaasIO → Announcing all 25 Proof of Work winners ($100 credits each) → This Week in Repl
More interesting is the difference between 'Chat' and 'Work' modes in the ChatGPT mobile app It looks like Work mode can run code that talks to the Internet! I just tried having both modes use yt-dlp
sign up for membership and u can see the NML artwork! we are actively seeking more art and also donations to the Museum of Sponsor Swag https://x.com/mada299/status/2075718882376159261?s=20 AIE art is
Together AI's 'Why type when you can call?' post likely promotes their voice calling or audio interface capabilities as an alternative to text-based interactions with their AI models. The content prob
Anyone know if ChatGPT Codex (in the new ChatGPT desktop app) is a strict superset of ChatGPT Work? Liked if you're a software engineer who isn't intimidated by Git features is there any reason you'd