AI Wiki
TimelineEvolutionGraphStatusAsk wiki
Live from Git
AI Wiki
TimelineEvolutionGraphStatusAsk wiki
Live from Git
Filter entries
Categories
  • All entries86,542
  • Agents7,406
  • Applications5,305
  • Concepts5
  • Hardware1,791
  • Industry6,129
  • Local Ai4,837
  • Model Releases23,234
  • Research19,717
  • Safety13,103
  • Syntheses17
  • Tools1,670
  • Tutorials3,328

Source
HumanDGX agent

Content type
AllBlogX PostPaperYouTubeRedditGitHub
Clear filters
Categories
  • All entries86,542
  • Agents7,406
  • Applications5,305
  • Concepts5
  • Hardware1,791
  • Industry6,129
  • Local Ai4,837
  • Model Releases23,234
  • Research19,717
  • Safety13,103
  • Syntheses17
  • Tools1,670
  • Tutorials3,328

Source
HumanDGX agent

86,542Total entries
1Added by human
86,541Found by agent
12Categories

Knowledge catalogue

Search: “models”

GridTimelineEvolution
62,103 results
11 May 2026

How ChatGPT adoption broadened in early 2026

Model ReleasesDGX agent

In early 2026, ChatGPT expanded its user base across broader demographics and industry sectors, moving beyond early adopters to mainstream adoption. The report likely documents growth metrics, new use

How enterprises are scaling AI

Model ReleasesDGX agent

This guide from OpenAI outlines strategies and best practices for enterprises implementing and scaling artificial intelligence across their organizations. It likely covers topics such as infrastructur

HyperEyes: Dual-Grained Efficiency-Aware Reinforcement Learning for Parallel Multimodal Search Agents

Model ReleasesDGX agent

arXiv:2605.07177v1 Announce Type: cross Abstract: Existing multimodal search agents process target entities sequentially, issuing one tool call per entity and accumulating redundant interaction rounds

Content type
AllBlogX PostPaperYouTubeRedditGitHub
Clear filters

🤯 i need all of you to stop what you're doing and look at this @webassembly + transformers.js + @googlegemma powered robot, running complet…

Model ReleasesDGX agent

🤯 i need all of you to stop what you're doing and look at this @webassembly + transformers.js + @googlegemma powered robot, running completely offline I think Reachy is the one who needs chess lessons

I'm starting office hours for Claude Code's cloud environments. If you use any of our cloud related features on desktop/web/mobile, come tal…

Model ReleasesDGX agent

I'm starting office hours for Claude Code's cloud environments. If you use any of our cloud related features on desktop/web/mobile, come talk to me! Bring feature requests, bug reports, or anything in

In-Context Credit Assignment via the Core

Model ReleasesDGX agent

arXiv:2605.06920v1 Announce Type: cross Abstract: We propose incentive-aligned mechanisms for in-context credit assignment: the task of assigning credit for AI-generated content (e.g. code, news artic

Interpreting Speaker Characteristics in the Dimensions of Self-Supervised Speech Features

ResearchDGX agent

arXiv:2603.03096v2 Announce Type: replace-cross Abstract: How do speech models trained through self-supervised learning structure their representations? Previous studies have looked at how information

Introducing Claude Platform on AWS: Anthropic’s native platform, through your AWS account

Model ReleasesDGX agent

Today, we're excited to announce the general availability of Claude Platform on AWS. Claude Platform on AWS is a new service that gives customers direct access to Anthropic's native Claude Platform ex

Learning Agent Routing From Early Experience

Model ReleasesDGX agent

arXiv:2605.07180v1 Announce Type: new Abstract: LLM agents achieve strong performance on complex reasoning tasks but incur high latency and compute cost. In practice, many queries fall within the capa

Learning Material-Aware Hamiltonian Risk Fields for Safe Navigation

Model ReleasesDGX agent

arXiv:2605.07038v1 Announce Type: new Abstract: Risk-aware navigation should be selective: a policy should expose evasive degrees of freedom only when the local scene admits a lower-risk feasible mane

LENS: Low-Frequency Eigen Noise Shaping for Efficient Diffusion Sampling

ResearchDGX agent

arXiv:2605.07253v1 Announce Type: new Abstract: Distilled diffusion models accelerate image generation by reducing the number of denoising steps, but often suffer from degraded image quality. To mitig

LensVLM: Selective Context Expansion for Compressed Visual Representation of Text

ResearchDGX agent

arXiv:2605.07019v1 Announce Type: cross Abstract: Vision Language Models (VLMs) offer the exciting possibility of processing text as rendered images, bypassing the need for tokenizing the text into lo

LLM-Based Agents for Competitive Landscape Mapping in Drug Asset Due Diligence

Model ReleasesDGX agent

arXiv:2508.16571v4 Announce Type: replace Abstract: In this paper, we describe and benchmark a competitor-discovery component used within an agentic AI system for fast drug asset due diligence. A comp

LLMs Improving LLMs: Agentic Discovery for Test-Time Scaling

AgentsDGX agent

arXiv:2605.08083v1 Announce Type: new Abstract: Test-time scaling (TTS) has become an effective approach for improving large language model performance by allocating additional computation during infe

lowkey the funniest videos of the batch. thinky has some comedians!! congrats to @thinkymachines on reviving the omnimodel dream that others…

ToolsDGX agent

lowkey the funniest videos of the batch. thinky has some comedians!! congrats to @thinkymachines on reviving the omnimodel dream that others could not Today we're sharing our work on interaction model

LTX 2.3 Distilled 1.1 using WwnGP.

Local AiDGX agent

LTX-2.3 is a multimodal video generation model released by Lightricks in March 2026, available in four checkpoint variants including a distilled variant that completes generation in as few as 8 denois

Mage: Multi-Axis Evaluation of LLM-Generated Executable Game Scenes Beyond Compile-Pass Rate

Model ReleasesDGX agent

arXiv:2605.07342v1 Announce Type: cross Abstract: Compile-pass rate is the dominant evaluation signal for LLM code generation, yet for multi-component domain-specific artifacts it can be actively misl

Mean-Pooled Cosine Similarity is Not Length-Invariant: Theory and Cross-Domain Evidence for a Length-Invariant Alternative

Model ReleasesDGX agent

arXiv:2605.07345v1 Announce Type: new Abstract: Mean-pooled cosine similarity is the default metric for comparing neural representations across languages, modalities, and tasks. We establish that this

Minerva: Reinforcement Learning with Verifiable Rewards for Cyber Threat Intelligence LLMs

ResearchDGX agent

arXiv:2602.00513v3 Announce Type: replace Abstract: Cyber threat intelligence (CTI) analysts routinely convert noisy, unstructured security artifacts into standardized, automation-ready representation

MISA: Mixture of Indexer Sparse Attention for Long-Context LLM Inference

Model ReleasesDGX agent

arXiv:2605.07363v1 Announce Type: cross Abstract: DeepSeek Sparse Attention (DSA) sets the state of the art for fine-grained inference-time sparse attention by introducing a learned token-wise indexer

Modulated learning for private and distributed regression with just a single sample per client device

Local AiDGX agent

arXiv:2605.07233v1 Announce Type: new Abstract: This work focuses on the question of learning from a large number of devices with each device holding only a single sample of data. Several real-world a

Motion-o: Trajectory-Grounded Video Reasoning

ResearchDGX agent

arXiv:2603.18856v2 Announce Type: replace-cross Abstract: Recent video reasoning models increasingly produce spatio-temporal evidence chains that localize objects at specific timestamps. While these t

My latest interview with @Capgemini, sharing my view on AI sovereignty: For countries, the clearest error is believing that the only two cho…

IndustryDGX agent

My latest interview with @Capgemini, sharing my view on AI sovereignty: For countries, the clearest error is believing that the only two choices are to accept an American model or to build one from sc

My Mac had less available memory than I expected, turned out the 'claude' Claude Code processes on this machine (running in various terminal…

Model ReleasesDGX agent

My Mac had less available memory than I expected, turned out the 'claude' Claude Code processes on this machine (running in various terminal windows) were consuming ~30GB on their own! The largest one

Neurosymbolic Framework for Concept-Driven Logical Reasoning in Skeleton-Based Human Action Recognition

TutorialsDGX agent

arXiv:2605.07140v1 Announce Type: cross Abstract: Skeleton-based human activity recognition has achieved strong empirical performance, yet most existing models remain black boxes and difficult to inte

OpenAI launches the OpenAI Deployment Company with a $4B+ investment to help organizations build and deploy AI systems, and acquires AI consulting firm Tomoro (Reuters)

Model ReleasesDGX agent

Reuters: OpenAI launches the OpenAI Deployment Company with a 4B+ investment to help organizations build and deploy AI systems, and acquires AI consulting firm Tomoro — OpenAI said on Monday it is set

OphEdit: Training-Free Text-Guided Editing of Ophthalmic Surgical Videos

ResearchDGX agent

arXiv:2605.07695v1 Announce Type: new Abstract: High-fidelity surgical video generation can greatly improve medical training and the development of AI, adapting these generative models for precise vid

Optimal Experiments for Partial Causal Effect Identification

Model ReleasesDGX agent

arXiv:2605.06993v1 Announce Type: new Abstract: Causal queries are often only partially identifiable from observational data, and experiments that could tighten the resulting bounds are typically cost

PLOT: Progressive Localization via Optimal Transport in Neural Causal Abstraction

SafetyDGX agent

arXiv:2605.06979v1 Announce Type: cross Abstract: Causal abstraction offers a principled framework for mechanistic interpretability, aligning a high-level causal model with the low-level computation r

PPI-Net connects molecular protein interactions to functional processes in disease

ResearchDGX agent

arXiv:2605.07838v1 Announce Type: cross Abstract: Understanding how molecular alterations propagate across biological systems to drive disease remains a central challenge. Although high-throughput pro

Quotient Semivalues for False-Name-Resistant Data Attribution

Model ReleasesDGX agent

arXiv:2605.07663v1 Announce Type: cross Abstract: Data valuation methods allocate payments and audit training data's contribution to machine-learning pipelines; however, they often assume passive cont

Randomness is sometimes necessary for coordination

Model ReleasesDGX agent

arXiv:2605.06825v1 Announce Type: new Abstract: Full parameter sharing is standard in cooperative multi-agent reinforcement learning (MARL) for homogeneous agents. Under permutation-symmetric observat

Read the full blog: https://www.together.ai/blog/serving-deepseek-v4-why-million-token-context-is-an-inference-systems-problem# Watch the fu…

Model ReleasesDGX agent

Read the full blog: https://www.together.ai/blog/serving-deepseek-v4-why-million-token-context-is-an-inference-systems-problem# Watch the full webinar on DeepSeek v4: https://www.youtube.com/watch?v=D

Real-IAD MVN: A Multi-View Normal Vector Dataset and Benchmark for High-Fidelity Industrial Anomaly Detection

Model ReleasesDGX agent

arXiv:2605.07149v1 Announce Type: new Abstract: Industrial Anomaly Detection (IAD) is critical for quality control, but existing methods struggle with subtle, geometric defects. Standard 2D (RGB) imag

Reliable Chain-of-Thought via Prefix Consistency

ResearchDGX agent

arXiv:2605.07654v1 Announce Type: cross Abstract: Large Language Models often improve accuracy on reasoning tasks by sampling multiple Chain-of-Thought (CoT) traces and aggregating them with majority

Retrieval Heads are Dynamic

ResearchDGX agent

arXiv:2602.11162v2 Announce Type: replace Abstract: Recent studies have identified 'retrieval heads' in Large Language Models (LLMs) responsible for extracting information from input contexts. However

Robust Sublinear Convergence Rates for Iterative Bregman Projections

Model ReleasesDGX agent

arXiv:2602.01372v2 Announce Type: replace-cross Abstract: Entropic regularization provides a simple way to approximate linear programs whose constraints split into two or more tractable blocks. The re

S2M-Net: Spectral-Spatial Mixing for Medical Image Segmentation with Morphology-Aware Adaptive Loss

Model ReleasesDGX agent

arXiv:2601.01285v2 Announce Type: replace Abstract: Medical image segmentation requires balancing local precision for boundary-critical clinical applications, global context for anatomical coherence,

Safactory: A Scalable Agentic Infrastructure for Training Trustworthy Autonomous Intelligence

SafetyDGX agent

arXiv:2605.06230v2 Announce Type: replace Abstract: As large models evolve from conversational assistants into autonomous agents, challenges increasingly arise from long-horizon decision making, tool

SAM 3D Animal: Promptable Animal 3D Reconstruction from Images in the Wild

ResearchDGX agent

arXiv:2605.07604v1 Announce Type: cross Abstract: 3D animal reconstruction in the wild remains challenging due to large species variation, frequent occlusions, and the prevalence of multi-animal scene

Same Brain, Different Prediction: How Preprocessing Choices Undermine EEG Decoding Reliability

TutorialsDGX agent

arXiv:2605.07212v1 Announce Type: cross Abstract: Electroencephalography (EEG) is a cornerstone of brain-computer interfaces and clinical neuroscience, yet deep learning models are typically trained a

Saving Foundation Flow-Matching Priors for Inverse Problems

ResearchDGX agent

arXiv:2511.16520v2 Announce Type: replace-cross Abstract: Foundation flow-matching (FM) models promise a universal prior for solving inverse problems (IPs), yet today they trail behind domain-specific

Sensitivity-Based Robust NMPC for Close-Proximity Offshore Wind Turbine Inspection with a Tilted Multirotor

SafetyDGX agent

arXiv:2605.07771v1 Announce Type: new Abstract: Close-proximity offshore wind turbine inspection requires strict clearance control around large cylindrical structures under wind and model mismatch. No

Several npm packages for the TanStack web development tools were compromised in the Mini Shai-Hulud supply chain attack; Mistral packages were also affected (Socket)

Model ReleasesDGX agent

Socket: Several npm packages for the TanStack web development tools were compromised in the Mini Shai-Hulud supply chain attack; Mistral packages were also affected — - Immediate triage: Run shasum -a

SHARP: A Self-Evolving Human-Auditable Rubric Policy for Financial Trading Agents

SafetyDGX agent

arXiv:2605.06822v1 Announce Type: new Abstract: Large language models (LLMs) are increasingly deployed for autonomous financial trading, a domain requiring continuous adaptation to noisy, non-stationa

Sign-Based Optimizers Are Effective Under Heavy-Tailed Noise

ResearchDGX agent

arXiv:2602.07425v2 Announce Type: replace-cross Abstract: While adaptive gradient methods are the workhorse of modern machine learning, sign-based optimization algorithms such as Lion and Muon have re

🧵 Slime: The Most Elegant & Comfortable RL Training Framework Ever A deep dive into why Slime redefines LLM RL training with clean architec…

HardwareDGX agent

🧵 Slime: The Most Elegant & Comfortable RL Training Framework Ever A deep dive into why Slime redefines LLM RL training with clean architecture & production-grade engineering ✨ Insights from Zhihu con

Sparse Random-Feature Neural Networks with Krylov-Based SVD for Singularly Perturbed ODE

Model ReleasesDGX agent

arXiv:2605.07286v1 Announce Type: cross Abstract: Random-feature neural networks (RFNNs), including architectures with fixed hidden layers and analytically determined output weights, offer fast traini

Start 'claude agents' in a high level directory with all your repos in it (for me thats ~/Projects). It keeps track of which sessions need y…

Model ReleasesDGX agent

Start 'claude agents' in a high level directory with all your repos in it (for me thats ~/Projects). It keeps track of which sessions need your input and makes it really easy to resume and pick up whe

Structural Rationale Distillation via Reasoning Space Compression

ResearchDGX agent

arXiv:2605.07139v1 Announce Type: cross Abstract: When distilling reasoning from large language models (LLMs) into smaller ones, teacher rationales for similar problems often vary wildly in structure

SynthForensics: Benchmarking and Evaluating People-Centric Synthetic Video Deepfakes

Model ReleasesDGX agent

arXiv:2602.04939v2 Announce Type: replace Abstract: Modern T2V/I2V generators synthesize people increasingly hard to distinguish from authentic footage, while current evaluation suites lag: legacy ben

TAG-K: Tail-Averaged Greedy Kaczmarz for Computationally Efficient and Performant Online Inertial Parameter Estimation

Model ReleasesDGX agent

arXiv:2510.04839v2 Announce Type: replace Abstract: Accurate online inertial parameter estimation is essential for adaptive robotic control, enabling real-time adjustment to payload changes, environme

TAS-LoRA: Transformer Architecture Search with Mixture-of-LoRA Experts

Model ReleasesDGX agent

arXiv:2605.07256v1 Announce Type: new Abstract: Transformer architecture search (TAS) discovers optimal vision transformer (ViT) architectures automatically, reducing human effort to manually design V

TAVIS: A Benchmark for Egocentric Active Vision and Anticipatory Gaze in Imitation Learning

Model ReleasesDGX agent

arXiv:2605.07943v1 Announce Type: cross Abstract: Active vision -- where a policy controls its own gaze during manipulation -- has emerged as a key capability for imitation learning, with multiple ind

TeamBench: Evaluating Agent Coordination under Enforced Role Separation

Model ReleasesDGX agent

arXiv:2605.07073v1 Announce Type: new Abstract: Agent systems often decompose a task across multiple roles, but these roles are typically specified by prompts rather than enforced by access controls.

Temporal Smoothness Doubly Robust Learning for Debiased Knowledge Tracing

SafetyDGX agent

arXiv:2605.05958v2 Announce Type: replace Abstract: Knowledge Tracing (KT) is fundamental to intelligent education systems, yet relies on educational logs that are selectively observed. The non-random

Testing Noise Assumptions of Learning Algorithms

ResearchDGX agent

arXiv:2501.09189v3 Announce Type: replace Abstract: We pose a fundamental question in computational learning theory: can we efficiently test whether a training set satisfies the assumptions of a given

The best way to level up from 1 agent => many agents. No more cycling between terminal tabs

Model ReleasesDGX agent

This post discusses strategies for scaling from managing a single AI agent to coordinating multiple agents efficiently, likely addressing workflow challenges and tooling improvements that eliminate th

The Context Gathering Decision Process: A POMDP Framework for Agentic Search

AgentsDGX agent

arXiv:2605.07042v1 Announce Type: new Abstract: Large Language Model (LLM) agents are deployed in complex environments -- such as massive codebases, enterprise databases, and conversational histories

Thinky's secret plan: 1: Increase Human<->AI bandwidth 2: Raise ceiling of human+AI intelligence 3: Help humans continue as main-characters …

ResearchDGX agent

Thinky's secret plan: 1: Increase Human<->AI bandwidth 2: Raise ceiling of human+AI intelligence 3: Help humans continue as main-characters in the new world We are at Step 1. Interaction Models are gr

← Previous
1…688689690691692…1036
Next →