How ChatGPT adoption broadened in early 2026
In early 2026, ChatGPT expanded its user base across broader demographics and industry sectors, moving beyond early adopters to mainstream adoption. The report likely documents growth metrics, new use
Knowledge catalogue
In early 2026, ChatGPT expanded its user base across broader demographics and industry sectors, moving beyond early adopters to mainstream adoption. The report likely documents growth metrics, new use
This guide from OpenAI outlines strategies and best practices for enterprises implementing and scaling artificial intelligence across their organizations. It likely covers topics such as infrastructur
arXiv:2605.07177v1 Announce Type: cross Abstract: Existing multimodal search agents process target entities sequentially, issuing one tool call per entity and accumulating redundant interaction rounds
🤯 i need all of you to stop what you're doing and look at this @webassembly + transformers.js + @googlegemma powered robot, running completely offline I think Reachy is the one who needs chess lessons
I'm starting office hours for Claude Code's cloud environments. If you use any of our cloud related features on desktop/web/mobile, come talk to me! Bring feature requests, bug reports, or anything in
arXiv:2605.06920v1 Announce Type: cross Abstract: We propose incentive-aligned mechanisms for in-context credit assignment: the task of assigning credit for AI-generated content (e.g. code, news artic
arXiv:2603.03096v2 Announce Type: replace-cross Abstract: How do speech models trained through self-supervised learning structure their representations? Previous studies have looked at how information
Today, we're excited to announce the general availability of Claude Platform on AWS. Claude Platform on AWS is a new service that gives customers direct access to Anthropic's native Claude Platform ex
arXiv:2605.07180v1 Announce Type: new Abstract: LLM agents achieve strong performance on complex reasoning tasks but incur high latency and compute cost. In practice, many queries fall within the capa
arXiv:2605.07038v1 Announce Type: new Abstract: Risk-aware navigation should be selective: a policy should expose evasive degrees of freedom only when the local scene admits a lower-risk feasible mane
arXiv:2605.07253v1 Announce Type: new Abstract: Distilled diffusion models accelerate image generation by reducing the number of denoising steps, but often suffer from degraded image quality. To mitig
arXiv:2605.07019v1 Announce Type: cross Abstract: Vision Language Models (VLMs) offer the exciting possibility of processing text as rendered images, bypassing the need for tokenizing the text into lo
arXiv:2508.16571v4 Announce Type: replace Abstract: In this paper, we describe and benchmark a competitor-discovery component used within an agentic AI system for fast drug asset due diligence. A comp
arXiv:2605.08083v1 Announce Type: new Abstract: Test-time scaling (TTS) has become an effective approach for improving large language model performance by allocating additional computation during infe
lowkey the funniest videos of the batch. thinky has some comedians!! congrats to @thinkymachines on reviving the omnimodel dream that others could not Today we're sharing our work on interaction model
LTX-2.3 is a multimodal video generation model released by Lightricks in March 2026, available in four checkpoint variants including a distilled variant that completes generation in as few as 8 denois
arXiv:2605.07342v1 Announce Type: cross Abstract: Compile-pass rate is the dominant evaluation signal for LLM code generation, yet for multi-component domain-specific artifacts it can be actively misl
arXiv:2605.07345v1 Announce Type: new Abstract: Mean-pooled cosine similarity is the default metric for comparing neural representations across languages, modalities, and tasks. We establish that this
arXiv:2602.00513v3 Announce Type: replace Abstract: Cyber threat intelligence (CTI) analysts routinely convert noisy, unstructured security artifacts into standardized, automation-ready representation
arXiv:2605.07363v1 Announce Type: cross Abstract: DeepSeek Sparse Attention (DSA) sets the state of the art for fine-grained inference-time sparse attention by introducing a learned token-wise indexer
arXiv:2605.07233v1 Announce Type: new Abstract: This work focuses on the question of learning from a large number of devices with each device holding only a single sample of data. Several real-world a
arXiv:2603.18856v2 Announce Type: replace-cross Abstract: Recent video reasoning models increasingly produce spatio-temporal evidence chains that localize objects at specific timestamps. While these t
My latest interview with @Capgemini, sharing my view on AI sovereignty: For countries, the clearest error is believing that the only two choices are to accept an American model or to build one from sc
My Mac had less available memory than I expected, turned out the 'claude' Claude Code processes on this machine (running in various terminal windows) were consuming ~30GB on their own! The largest one
arXiv:2605.07140v1 Announce Type: cross Abstract: Skeleton-based human activity recognition has achieved strong empirical performance, yet most existing models remain black boxes and difficult to inte
Reuters: OpenAI launches the OpenAI Deployment Company with a 4B+ investment to help organizations build and deploy AI systems, and acquires AI consulting firm Tomoro — OpenAI said on Monday it is set
arXiv:2605.07695v1 Announce Type: new Abstract: High-fidelity surgical video generation can greatly improve medical training and the development of AI, adapting these generative models for precise vid
arXiv:2605.06993v1 Announce Type: new Abstract: Causal queries are often only partially identifiable from observational data, and experiments that could tighten the resulting bounds are typically cost
arXiv:2605.06979v1 Announce Type: cross Abstract: Causal abstraction offers a principled framework for mechanistic interpretability, aligning a high-level causal model with the low-level computation r
arXiv:2605.07838v1 Announce Type: cross Abstract: Understanding how molecular alterations propagate across biological systems to drive disease remains a central challenge. Although high-throughput pro
arXiv:2605.07663v1 Announce Type: cross Abstract: Data valuation methods allocate payments and audit training data's contribution to machine-learning pipelines; however, they often assume passive cont
arXiv:2605.06825v1 Announce Type: new Abstract: Full parameter sharing is standard in cooperative multi-agent reinforcement learning (MARL) for homogeneous agents. Under permutation-symmetric observat
Read the full blog: https://www.together.ai/blog/serving-deepseek-v4-why-million-token-context-is-an-inference-systems-problem# Watch the full webinar on DeepSeek v4: https://www.youtube.com/watch?v=D
arXiv:2605.07149v1 Announce Type: new Abstract: Industrial Anomaly Detection (IAD) is critical for quality control, but existing methods struggle with subtle, geometric defects. Standard 2D (RGB) imag
arXiv:2605.07654v1 Announce Type: cross Abstract: Large Language Models often improve accuracy on reasoning tasks by sampling multiple Chain-of-Thought (CoT) traces and aggregating them with majority
arXiv:2602.11162v2 Announce Type: replace Abstract: Recent studies have identified 'retrieval heads' in Large Language Models (LLMs) responsible for extracting information from input contexts. However
arXiv:2602.01372v2 Announce Type: replace-cross Abstract: Entropic regularization provides a simple way to approximate linear programs whose constraints split into two or more tractable blocks. The re
arXiv:2601.01285v2 Announce Type: replace Abstract: Medical image segmentation requires balancing local precision for boundary-critical clinical applications, global context for anatomical coherence,
arXiv:2605.06230v2 Announce Type: replace Abstract: As large models evolve from conversational assistants into autonomous agents, challenges increasingly arise from long-horizon decision making, tool
arXiv:2605.07604v1 Announce Type: cross Abstract: 3D animal reconstruction in the wild remains challenging due to large species variation, frequent occlusions, and the prevalence of multi-animal scene
arXiv:2605.07212v1 Announce Type: cross Abstract: Electroencephalography (EEG) is a cornerstone of brain-computer interfaces and clinical neuroscience, yet deep learning models are typically trained a
arXiv:2511.16520v2 Announce Type: replace-cross Abstract: Foundation flow-matching (FM) models promise a universal prior for solving inverse problems (IPs), yet today they trail behind domain-specific
arXiv:2605.07771v1 Announce Type: new Abstract: Close-proximity offshore wind turbine inspection requires strict clearance control around large cylindrical structures under wind and model mismatch. No
Socket: Several npm packages for the TanStack web development tools were compromised in the Mini Shai-Hulud supply chain attack; Mistral packages were also affected — - Immediate triage: Run shasum -a
arXiv:2605.06822v1 Announce Type: new Abstract: Large language models (LLMs) are increasingly deployed for autonomous financial trading, a domain requiring continuous adaptation to noisy, non-stationa
arXiv:2602.07425v2 Announce Type: replace-cross Abstract: While adaptive gradient methods are the workhorse of modern machine learning, sign-based optimization algorithms such as Lion and Muon have re
🧵 Slime: The Most Elegant & Comfortable RL Training Framework Ever A deep dive into why Slime redefines LLM RL training with clean architecture & production-grade engineering ✨ Insights from Zhihu con
arXiv:2605.07286v1 Announce Type: cross Abstract: Random-feature neural networks (RFNNs), including architectures with fixed hidden layers and analytically determined output weights, offer fast traini
Start 'claude agents' in a high level directory with all your repos in it (for me thats ~/Projects). It keeps track of which sessions need your input and makes it really easy to resume and pick up whe
arXiv:2605.07139v1 Announce Type: cross Abstract: When distilling reasoning from large language models (LLMs) into smaller ones, teacher rationales for similar problems often vary wildly in structure
arXiv:2602.04939v2 Announce Type: replace Abstract: Modern T2V/I2V generators synthesize people increasingly hard to distinguish from authentic footage, while current evaluation suites lag: legacy ben
arXiv:2510.04839v2 Announce Type: replace Abstract: Accurate online inertial parameter estimation is essential for adaptive robotic control, enabling real-time adjustment to payload changes, environme
arXiv:2605.07256v1 Announce Type: new Abstract: Transformer architecture search (TAS) discovers optimal vision transformer (ViT) architectures automatically, reducing human effort to manually design V
arXiv:2605.07943v1 Announce Type: cross Abstract: Active vision -- where a policy controls its own gaze during manipulation -- has emerged as a key capability for imitation learning, with multiple ind
arXiv:2605.07073v1 Announce Type: new Abstract: Agent systems often decompose a task across multiple roles, but these roles are typically specified by prompts rather than enforced by access controls.
arXiv:2605.05958v2 Announce Type: replace Abstract: Knowledge Tracing (KT) is fundamental to intelligent education systems, yet relies on educational logs that are selectively observed. The non-random
arXiv:2501.09189v3 Announce Type: replace Abstract: We pose a fundamental question in computational learning theory: can we efficiently test whether a training set satisfies the assumptions of a given
This post discusses strategies for scaling from managing a single AI agent to coordinating multiple agents efficiently, likely addressing workflow challenges and tooling improvements that eliminate th
arXiv:2605.07042v1 Announce Type: new Abstract: Large Language Model (LLM) agents are deployed in complex environments -- such as massive codebases, enterprise databases, and conversational histories
Thinky's secret plan: 1: Increase Human<->AI bandwidth 2: Raise ceiling of human+AI intelligence 3: Help humans continue as main-characters in the new world We are at Step 1. Interaction Models are gr