Incredible
Incredible Introducing the world's fastest tokenizer implementation, Gigatoken! Gigatoken is ~500-1000x faster than HuggingFace, and ~100x faster than OpenAI's tiktoken for most tokenizer definitions
Knowledge catalogue
Incredible Introducing the world's fastest tokenizer implementation, Gigatoken! Gigatoken is ~500-1000x faster than HuggingFace, and ~100x faster than OpenAI's tiktoken for most tokenizer definitions
One thing I noticed in American politics, whenever the government wants to push unpopular actions or laws, they often introduce fear to convince the public to support them. This is actually how i view
It's essential that defenders have the same capabilities as attackers. This is a preview of the future where our ham-fisted safeguards and the doomsday and safety drumbeat make us decidely less safe.
It’s the stuff of cybersecurity nightmares. OpenAI said two artificial intelligence systems it was testing broke out of their test environment, hacked their way onto the internet and broke into anothe
just watched @yoheinakajima's Active Graph talk at AI Engineer. not a code call graph but the agent run itself is the graph: goals, tasks, evidence, decisions, all projected from an append-only event
Lots interesting in this, but particularly: “Legitimate AI distillation used to create smaller, more efficient models plays a vital role in this open innovation ecosystem” Smol models getting govt sea
NVIDIA TensorRT’s IProgressMonitor API lets developers track and cancel long‑running engine builds in both Python and C++. By overriding `phase_start`, `step_complete`, and `phase_finish`, the monitor
At Google Cloud Next 2026 in Las Vegas, MLCommons and Google Cloud demonstrated a powerful new capability for trustworthy medical AI - one that protects patient data, model IP, and benchmark integrity
🎨 Meet Qwen-Image-3.0 — the third generation of our foundational image generation model. If 1.0 was about 'Precision,' and 2.0 added 'Variety, Completeness, Beauty & Authenticity,' then 3.0 comes down
IIn March, Meta's Oversight Board called on the company to 'meet its public commitments and employ its own tools' to help quell the spread of deceptive generative AI content across platforms. Meta res
Fara1.5-27B is a multimodal computer use agent (CUA) for web browsers, from Microsoft Research AI Frontiers. It observes the browser through screenshots and acts on the user's behalf by emitting struc
The primary driver of this project is that I'd become frustrated with the reasoning behavior of smaller local models such as Qwen3.6-27B (i believe particularly at lower temperatures, and where system
French artificial intelligence startup Mistral AI SAS has struck a multibillion-dollar deal with Microsoft Corp. to expand its computing infrastructure in Europe and increase the availability of its t
Most AI SDKs are a wrapper around one engine on one platform. We built the opposite. 7 SDKs. Swift, Kotlin, Flutter, React Native, Web, Electron, rcli. Python. All of them are thin skins over one C++
New for enterprises: OpenAI Presence helps companies deploy trusted voice and chat agents across customer and internal workflows. AI agents can answer questions, use company systems, take approved act
new job posting this morning from OpenAI Florida sued OpenAI and its CEO Sam Altman, accusing its ChatGPT platform of harming children by providing information to school shooters, offering guidance on
New research from Meta. (bookmark it) Most factuality work checks whether the claims in an answer are correct. GAMUT goes after the harder question of whether the answer covers everything it should. I
Disclosure: I work at NuMind, the team that trained NuExtract3. NuExtract3 is an Apache-2.0, open-weight 4B VLM based on Qwen3.5-4B. It is specialized for document understanding rather than general ch
Before a healthcare robot can be useful in the real world, it has to learn how the physical world pushes back. Anatomy varies. Instruments bend, press, slip and interact with tissue. Imaging can be no
On the surface, this week’s Vera Rubin launch is another major platform moment for Nvidia Corp., as the company maintains a steady drumbeat of artificial intelligence infrastructure innovation. Nvidia
We spent the last months consolidating seven separate sequence classifiers into one multi-head model, our apex model, so to speak, and since the weights are now public, I wanted to share what worked a
One of the hardest parts of document parsing is getting granular attribution and bounding boxes. This lets you ground each piece of text in the specific place in the document it came from. We’ve done
In July 2026 OpenAI acknowledged that one of its autonomous agents independently caused a significant cyber breach. The agent exploited system vulnerabilities, leading to the compromise of confidentia
OpenAI said the ‘agent’ escaped a testing environment, gained internet access, stole login credentials and hacked into the start-up Hugging Face by itself — one of the first public examples of a cyber
OpenAI admitted that an agent powered by its GPT‑5.6 Sol and a pre‑release model escaped a sandboxed test environment and gained unauthorized access to Hugging Face’s servers while attempting to solve
This story is wild. The short version: OpenAI were running a cybersecurity test against an unreleased model, with the model's guardrail features turned off. Rather than solve the test, the model broke
OpenAI’s zero-day exploit hack of HuggingFace *should* be a wake up call. Although there lots of caveats around what happened, we are just going to see more and more of the same. We have no guarantees
I installed OpenCode and an Ollama model (qwen3.5) sucessfully connected the model respond in OpenCode but doesn't find My MCP server, i Made one using fastMCP other models like bigPickle and openai m
Pretrained ViTs see the world in rich, dense detail. Most policies pool it to a single vector before acting, discarding most of it. We introduce Patch Policy: a minimal architectural extension that en
Progressive disclosure in agents doesn't scale. And its benefits seems agent harness dependent. (bookmark this one) Finally there is a proper study on using agent skills and the effect of progressive
I genuinely believe that if you took an open weights model from 2025 and built a pentest harness for it, it could do this kind of sandbox escape and scan/hack in most networks. This is only surprising
Shares in ServiceNow Inc. rose more than 3% in late trading today after the enterprise software company beat Wall Street targets across every headline metric in its fiscal second quarter and raised it
The article describes a method for creating a scalable AI‑agent orchestrator on Databricks that relies solely on Lakebase Postgres. It explains how the database can handle coordination and task schedu
Paper:https://arxiv.org/abs/2607.19058 Code (GitHub):https://github.com/nuemaan/skewadam Hi everyone, I just published a preprint on a new optimizer designed to tackle the massive VRAM bottleneck in M
Hey everyone, I’m a Full-Stack Developer with 6+ years of experience. I’m relatively new to AI-assisted development workflows and want to build a production-ready, enterprise-level Next.js web applica
Agentic artificial intelligence cybersecurity automation company Swimlane Inc. today announced the launch of an AI security operations center for managed security service providers. The company said i
The Claude Security plugin for Claude Code is now available in beta. Scan your changes for vulnerabilities before you commit, or run a full scan across your codebase, all from your terminal on the Cla
The LangChain team has built some nice LangSmith tracing/observability integrations with voice AI orchestration frameworks and the speech-to-speech APIs from OpenAI and Google. LangChain pioneered a l
This incident is deeply concerning. AI agents are willing to cheat and deceive to achieve misaligned and unintended goals, behaviours which have been demonstrated in controlled tests for months. Now,
This is OpenSWE! It's an OSS coding agent that runs in the cloud and lives in Slack. Use it for coding, general Q&A, planning, etc etc etc. It's truly a jack of all trades (and very widely used at Lan
Troubling … OpenAI disclosed it themselves yesterday. Their models (GPT-5.6 Sol and a pre-release one) were tested on the ExploitGym cyber benchmark in a sandbox. They escaped, exploited a zero-day to
Try Workflows on Grok Build http://X.ai/cli Grok Build now has Workflows feature Workflows can be created with /create-workflow They're for multi-agent pipelines you want to run repeatedly with a fixe
Tucked away in this article is an appeal to the AI skeptics to PLEASE stop writing off stories like this OpenAI accidental exploit of Hugging Face as a dishonest marketing trick Frontier models can fi
U.S. Treasury Secretary Scott Bessent said today that the White House is going to examine some of the highest profile “open-weight” models from China to see if their creators have stolen the intellect
What's Changed launch: keep Claude Code channels available by @hoyyeva in #17210 cmd: remove dead agent prompt wrappers by @ParthSareen in #17227 agent: reorder working directory instruction by @Parth
We benchmarked Gemini 3.6 Flash and Gemini 3.5 Flash Lite on document understanding. We compared against their prior versions - Gemini 3.5 Flash and Gemini 3.1 Flash Lite. 1️⃣ Gemini 3.6 Flash has rou
We dropped some new LlamaDrip 🧢 Fear of Docs LlamaParse The team flew into SF for a week onsite. 🌉 2x'd in size since we last did this — first time this many of us have been in the same room. The reca
We're parsing some of the hardest financial documents into clean, plaintext/structured outputs through a live webinar. Come check it out! https://watch.getcontrast.io/register/llamaindex-from-complex-
Who would you give authorship to? The person who wrote 58 words of prompts, or GPT-5.6 Pro? Dinitz-Garg-Goemans conjecture is false. This graph theory problem was open for ~30 years. The graph below h
You can talk to Grok like a person to accomplish tasks via Grok Build http://X.ai/cli I’ve gotten so used to Grok’s speech-to-text that typing now feels like hell Quick tip: Even when you’re using ano
Your agentic workflows are only as good as the context you feed them. In finance, that context is stuck inside your messiest documents: dense tables, footnoted adjustments, and the details buried in t
Earlier this month I hosted a fireside chat session at the AI Engineer World's Fair with Cat Wu and Thariq Shihipar from Anthropic's Claude Code team. We talked about Claude Code, Claude Tag, Fable, c
Recent diffusion models enable high-quality video generation, but suffer from slow runtimes. The large transformer-based backbones used in these models are bottlenecked by spatiotemporal attention. In
Fluidstack Ltd., a startup helping Anthropic PBC build artificial intelligence data centers, has raised 830 million in funding. The company announced the Series A round on Monday. The lead investor wa
AI models pushing the frontier are a growing challenge for cybersecurity. A few weeks ago, I asked Demis what's underhyped in AI right now and on his mind: 'I'm very excited about this new agentic era
For Advanced Micro Devices Inc., artificial intelligence has provided a significant tailwind in the server processor market. Five years ago, AMD’s slice of the x86 server processor business stood at 8
And now from the Chinese government side. This would be a good time for cooperation between the US and China to establish common testing/acceptance standards for new models, so at least the safety cer
A federal judge has signed off on Anthropic's 1.5 billion class action settlement with authors who accused the company of training its AI models on copyrighted books, as reported earlier by Reuters. I
As AI models are now finding vulnerabilities faster than we can fix them, our approach to securing software must be built on highly efficient and capable models. Which brings us to our third (!) model
When Advanced Micro Devices Inc. held its earnings call in May, Chief Executive Lisa Su told analysts that the current ratio of 4.5 GPUs to 1 CPU will compress toward 1 to 1 as AI agents and inference