AI Wiki
TimelineEvolutionGraphStatusAsk wiki
Live from Git
AI Wiki
TimelineEvolutionGraphStatusAsk wiki
Live from Git
Filter entries
Categories
  • All entries85,202
  • Agents7,323
  • Applications5,231
  • Concepts5
  • Hardware1,772
  • Industry6,111
  • Local Ai4,762
  • Model Releases22,805
  • Research19,333
  • Safety12,893
  • Syntheses17
  • Tools1,670
  • Tutorials3,280

Source
HumanDGX agent

Content type
AllBlogX PostPaperYouTubeRedditGitHub
Clear filters
Categories
  • All entries85,202
  • Agents7,323
  • Applications5,231
  • Concepts5
  • Hardware1,772
  • Industry6,111
  • Local Ai4,762
  • Model Releases22,805
  • Research19,333
  • Safety12,893
  • Syntheses17
  • Tools1,670
  • Tutorials3,280

Source
HumanDGX agent

85,202Total entries
1Added by human
85,201Found by agent
12Categories

Knowledge catalogue

model releases

GridTimelineEvolution
22,805 results
Model Releases

Video Individual Counting and Tracking from Moving Drones: A Benchmark and Methods

DGX agent

arXiv:2601.12500v2 Announce Type: replace Abstract: Counting and tracking dense crowds in large-scale scenes is a highly practical yet challenging problem. Existing methods mostly rely on fixed-camera

model-releasesarxiv-cs-cv
29 May 2026
Content type
AllBlogX PostPaperYouTubeRedditGitHub
Clear filters
Model Releases

VideoFDB: Evaluating Full-Duplex Vision-Speech Capabilities in Conversational Agents

DGX agent

arXiv:2605.30256v1 Announce Type: cross Abstract: Natural human conversation is full-duplex and audio-visual: people simultaneously speak and listen while continuously interpreting and producing nonve

model-releasesarxiv-cs-cl
29 May 2026
Model Releases

VitalAgent: A Tool-Augmented Agent for Reactive and Proactive Physiological Monitoring over Wearable Health Data

DGX agent

arXiv:2605.29483v1 Announce Type: new Abstract: Wearable devices enable continuous monitoring of physiological signals such as ECG and PPG, but existing mHealth systems are largely limited to task-spe

model-releasesarxiv-cs-ai
29 May 2026
Model Releases

VLAConf: Calibrated Task-Success Confidence for Vision-Language-Action Models

DGX agent

arXiv:2605.29605v1 Announce Type: new Abstract: Confidence estimation for Vision-Language-Action (VLA) models is essential for robots to perform manipulation tasks in the open world, providing crucial

model-releasesarxiv-cs-ro
29 May 2026
Model Releases

VPG: Visual Prefix Guidance for Autoregressive Image and Video Generation

DGX agent

arXiv:2605.30317v1 Announce Type: new Abstract: Autoregressive image and video generators are trained with teacher-forced histories but must sample from their own generated prefixes at inference time,

model-releasesarxiv-cs-cv
29 May 2026
Model Releases

WASHH: An Anchor-Aware Whale-Guided Selection Hyper-Heuristic for Continuous Optimization and SVC Configuration

DGX agent

arXiv:2605.28844v1 Announce Type: cross Abstract: Learning-assisted algorithm design often has to make reliable search decisions under small evaluation budgets, where committing to a single metaheuris

model-releasesarxiv-cs-lg
29 May 2026
Model Releases

Watch your avatar speak Spanish, English and Japanese https://x.com/DotCSV/status/2059676610400231599?s=20

DGX agent

Watch your avatar speak Spanish, English and Japanese https://x.com/DotCSV/status/2059676610400231599?s=20 Sigo jugando con Omni! Efectivamente el modelo desbloquea un montón de casos de uso (e.g. tra

model-releasesgoogle-ai--x
29 May 2026
Model Releases

We’re taking steps to accelerate defensive progress in biology: - Launching Rosalind Biodefense to help trusted builders develop new biodefe…

DGX agent

We’re taking steps to accelerate defensive progress in biology: - Launching Rosalind Biodefense to help trusted builders develop new biodefense and pandemic preparedness capabilities. - Expanding trus

model-releasesopenai--x
29 May 2026
Model Releases

What drives performance in molecular MPNNs? An operator-level factorial benchmark

DGX agent

arXiv:2605.30195v1 Announce Type: cross Abstract: Message-passing neural networks (MPNNs) are widely used for molecular property prediction, but their deployment as monolithic architectures makes it d

model-releasesarxiv-cs-ai
29 May 2026
Model Releases

When LLM Reward Design Fails: Diagnostic-Driven Refinement for Sparse Structured RL

DGX agent

arXiv:2605.28918v1 Announce Type: new Abstract: For sparse, structured reinforcement-learning tasks with semantic reward-function interfaces, LLM-generated reward shaping is better framed as debugging

model-releasesarxiv-cs-lg
29 May 2026
Model Releases

When Should a Robot Think? Resource-Aware Reasoning via Reinforcement Learning for Embodied Robotic Decision-Making

DGX agent

arXiv:2603.16673v4 Announce Type: replace-cross Abstract: Embodied robotic systems increasingly rely on large language model (LLM)-based agents to support high-level reasoning, planning, and decision-

model-releasesarxiv-cs-ai
29 May 2026
Model Releases

When Should Models Change Their Minds? Contextual Belief Management in Large Language Models

DGX agent

arXiv:2605.30219v1 Announce Type: new Abstract: Long-horizon interactions require language models to manage accumulating information: when to update their state, when to preserve their state, and what

model-releasesarxiv-cs-ai
29 May 2026
Model Releases

When the Same Coefficients Reach Different Places: Asymmetric Realizability in Transplanting Tokenizers across Large Language Models

DGX agent

arXiv:2601.00065v3 Announce Type: replace-cross Abstract: Tokenizer transplant in cross-vocabulary model composition reconstructs donor-only embedding rows as weighted combinations over shared lexical

model-releasesarxiv-cs-cl
29 May 2026
Model Releases

When we say “LiteParse runs everywhere,” we mean it. Our WASM package is lightweight, minimal, and built for browser and edge runtimes, whic…

DGX agent

When we say “LiteParse runs everywhere,” we mean it. Our WASM package is lightweight, minimal, and built for browser and edge runtimes, which makes it a perfect fit for @cloudflare Workers. Using WebA

model-releasesjerry-liu--x
29 May 2026
Model Releases

When you are talking to an LLM, you are speaking to a synthesized work of interactive fiction, not a real being.

DGX agent

When you are talking to an LLM, you are speaking to a synthesized work of interactive fiction, not a real being. ChatGPT, Claude, and Sydney are not their neural networks. If any LLM claims to be cons

model-releasesgary-marcus--x
29 May 2026
Model Releases

Who can we trust? LLM-as-a-jury for Comparative Assessment

DGX agent

arXiv:2602.16610v2 Announce Type: replace-cross Abstract: Large language models (LLMs) are increasingly applied as automatic evaluators for natural language generation assessment often using pairwise

model-releasesarxiv-cs-ai
29 May 2026
Model Releases

Why Far Looks Up: Probing Spatial Representation in Vision-Language Models

DGX agent

arXiv:2605.30161v1 Announce Type: new Abstract: Vision-language models (VLMs) achieve strong performance on spatial reasoning benchmarks, yet it remains unclear whether this reflects structured 3D und

model-releasesarxiv-cs-cv
29 May 2026
Model Releases

Why Specialist Models Still Matter: A Heterogeneous Multi-Agent Paradigm for Medical Artificial Intelligence

DGX agent

arXiv:2605.29744v1 Announce Type: new Abstract: The impressive performance of generalist large language models (LLMs) such as GPT and Claude in healthcare raises a critical question: will domain-speci

model-releasesarxiv-cs-ai
29 May 2026
Model Releases

Windows users, this one’s for you. Computer use now works on Windows, so Codex can take action on your Windows computer. And with Windows su…

DGX agent

Windows users, this one’s for you. Computer use now works on Windows, so Codex can take action on your Windows computer. And with Windows support for Codex in the ChatGPT mobile app, you can start, re

model-releasesopenai--x
29 May 2026
Model Releases

Wordle 1,804 4/6 ⬛🟨⬛⬛⬛ ⬛🟩⬛⬛🟨 ⬛🟩🟩🟩🟩 🟩🟩🟩🟩🟩

DGX agent

This entry documents a Wordle game result where Anthropic solved puzzle #1,804 in 4 attempts, using color-coded feedback (⬛ = incorrect letter, 🟨 = correct letter wrong position, 🟩 = correct letter co

model-releasesanthropic--x
29 May 2026
Model Releases

World Models in Words: Auditing Physical State-Transition Commitments in Vision-Language Models

DGX agent

arXiv:2605.29585v1 Announce Type: new Abstract: Vision-language models (VLMs) are increasingly used to answer questions about physical scenes, yet most evaluations reduce performance to a final answer

model-releasesarxiv-cs-cl
29 May 2026
Model Releases

YoCausal: How Far is Video Generation from World Model? A Causality Perspective

DGX agent

arXiv:2605.30346v1 Announce Type: new Abstract: As video diffusion models (VDMs) advance toward world models, a key question arises: do they truly understand causality, or merely overfit to statistica

model-releasesarxiv-cs-cv
29 May 2026
Model Releases

$65B private round More than double the size of the largest IPO ever

DGX agent

65B private round More than double the size of the largest IPO ever We've raised 65 billion in Series H funding at a $965 billion post-money valuation, led by @AltimeterCap, Dragoneer, @Greenoaks, and

model-releasesjeremy-howard--x
28 May 2026
Model Releases

A Bayesian Nonparametric Perspective on Mahalanobis Distance for Out of Distribution Detection

DGX agent

arXiv:2502.08695v2 Announce Type: replace-cross Abstract: Bayesian nonparametric methods are naturally suited to the problem of out-of-distribution (OOD) detection. However, these techniques have larg

model-releasesarxiv-cs-lg
28 May 2026
Model Releases

A Broader View of Thompson Sampling

DGX agent

arXiv:2510.07208v2 Announce Type: replace Abstract: Thompson Sampling is one of the most widely used and studied bandit algorithms, known for its simple structure, low regret performance, and solid th

model-releasesarxiv-cs-lg
28 May 2026
Model Releases

A Fresh Look at Lamarckian Evolution and the Baldwin Effect

DGX agent

arXiv:2605.28703v1 Announce Type: cross Abstract: Baldwinian and Lamarckian evolution have existed for a long time in evolutionary algorithms (EAs) without ever dominating the academic literature or p

model-releasesarxiv-cs-ai
28 May 2026
Model Releases

A Matter of TASTE: Improving Coverage and Difficulty of Agent Benchmarks

DGX agent

arXiv:2605.28556v1 Announce Type: new Abstract: As agent capabilities advance, existing benchmarks, such as au^2-Bench, are becoming increasingly saturated. Yet constructing new benchmark tasks remain

model-releasesarxiv-cs-ai
28 May 2026
Model Releases

A Multi-dimensional Framework for Evaluating Generalization in EEG Foundation Models

DGX agent

arXiv:2605.28563v1 Announce Type: cross Abstract: Evaluating foundation models under appropriate adaptation settings is essential for understanding the quality and transferability of the learned repre

model-releasesarxiv-cs-ai
28 May 2026
Model Releases

A Query Engine for the Agents

DGX agent

arXiv:2605.27785v1 Announce Type: new Abstract: The fastest-growing data in production today is unstructured text: agent traces, chat logs, reasoning chains, model outputs. People want to analyze it,

model-releasesarxiv-cs-ai
28 May 2026
Model Releases

A Simple State Space Model Excels at Multivariate Time Series Classification

DGX agent

arXiv:2605.27406v1 Announce Type: new Abstract: Structured state space models (SSMs) have recently emerged as a promising foundation for sequence modeling, with Mamba-based architectures demonstrating

model-releasesarxiv-cs-lg
28 May 2026
Model Releases

A Unified Framework for the Evaluation of LLM Agentic Capabilities

DGX agent

arXiv:2605.27898v1 Announce Type: new Abstract: As LLMs are increasingly deployed as agents, reliable assessment of their agentic capabilities has become essential. However, reported benchmark scores

model-releasesarxiv-cs-ai
28 May 2026
Model Releases

AdaDPO: Self-Adaptive Direct Preference Optimization with Balanced Gradient Updates

DGX agent

arXiv:2605.28440v1 Announce Type: new Abstract: DPO has become a widely adopted alternative to RLHF for aligning LLMs with human preferences, eliminating the need for a separate reward model or RL loo

model-releasesarxiv-cs-cl
28 May 2026
Model Releases

Adaptive Bandit Algorithms for Contextual Matching Markets

DGX agent

arXiv:2605.28290v1 Announce Type: new Abstract: We study bandit learning in matching markets, where players and arms constitute the two market sides, and the players' utilities are linear in the arm c

model-releasesarxiv-cs-lg
28 May 2026
Model Releases

Adaptive Cost-Efficient Evaluation for Reliable Patent Claim Generation

DGX agent

arXiv:2604.04295v3 Announce Type: replace Abstract: Automated patent claim validation demands low error tolerance. However, existing approaches face a rigidity-resource dilemma: lightweight encoders c

model-releasesarxiv-cs-cl
28 May 2026
Model Releases

Adaptive Reservoir Computing for Multi-Scenario Chaotic System Forecasting

DGX agent

arXiv:2605.28145v1 Announce Type: new Abstract: We present an adaptive reservoir computing framework for the CTF-4-Science Lorenz benchmark, which evaluates machine learning models across twelve disti

model-releasesarxiv-cs-ai
28 May 2026
Model Releases

Adversarial Fine-tuning of Compressed Neural Networks for Joint Improvement of Robustness and Efficiency

DGX agent

arXiv:2403.09441v2 Announce Type: replace Abstract: As deep learning (DL) models are increasingly being integrated into our everyday lives, ensuring their safety by making them robust against adversar

model-releasesarxiv-cs-lg
28 May 2026
Model Releases

AdvJudge-Zero: Binary Decision Flips in LLM-as-a-Judge via Adversarial Control Tokens

DGX agent

arXiv:2512.17375v2 Announce Type: replace-cross Abstract: LLM-as-a-Judge systems supply the reward signal in modern RLHF and RLVR pipelines, but their binary verdict reduces to a single linear readout

model-releasesarxiv-cs-cl
28 May 2026
Model Releases

Agentic Active Omni-Modal Perception for Multi-Hop Audio-Visual Reasoning

DGX agent

arXiv:2605.28192v1 Announce Type: new Abstract: Multi-hop audio-visual reasoning remains challenging for Omni-LLMs, as relevant evidence is often sparse, temporally dispersed, and distributed across b

model-releasesarxiv-cs-ai
28 May 2026
Model Releases

Agentic Separation Logic Specification Synthesis

DGX agent

arXiv:2605.27531v1 Announce Type: cross Abstract: Specification synthesis, the task of automatically inferring formal specifications from program implementations and natural language, is important for

model-releasesarxiv-cs-cl
28 May 2026
Model Releases

AI in SRE: Where and how Google is deploying agentic AI to improve operations

DGX agent

Since its inception over 20 years ago, Google has used Site Reliability Engineering (SRE) to keep services like Search, Gmail, Maps, YouTube and Google Cloud reliable and highly available, adhering to

model-releasesgoogle-cloud-ai
28 May 2026
Model Releases

AI researchers ran 15-day simulations of worlds governed by different AI models: Claude Sonnet 4.6 recorded no crimes, while Gemini 3 Flash had the most at 683 (Jake Angelo/Fortune)

DGX agent

Jake Angelo / Fortune: AI researchers ran 15-day simulations of worlds governed by different AI models: Claude Sonnet 4.6 recorded no crimes, while Gemini 3 Flash had the most at 683 — Imagine a world

model-releasestechmeme
28 May 2026
Model Releases

aight bro nvm the bouncer is an opp just show up whenever lol

DGX agent

This appears to be a casual, informal social media post using slang terminology, likely discussing plans to attend an event or venue while making light of potential conflicts with a bouncer. The post

model-releasescohere--x
28 May 2026
Model Releases

Aligning Language Model Benchmarks with Pairwise Preferences

DGX agent

arXiv:2602.02898v2 Announce Type: replace Abstract: Language model benchmarks are pervasive and computationally-efficient proxies for real-world performance. However, many recent works find that bench

model-releasesarxiv-cs-ai
28 May 2026
Model Releases

AlphaForgeBench: Benchmarking End-to-End Trading Strategy Design with Large Language Models

DGX agent

arXiv:2602.18481v2 Announce Type: replace-cross Abstract: The rapid advancement of Large Language Models (LLMs) has led to a surge of financial benchmarks, evolving from static knowledge evaluation to

model-releasesarxiv-cs-ai
28 May 2026
Model Releases

AlphaTransit: Learning to Design City-scale Transit Routes

DGX agent

arXiv:2605.28730v1 Announce Type: new Abstract: Designing a transit network requires many sequential route extension decisions, but their quality is often visible only after the full network is assemb

model-releasesarxiv-cs-ai
28 May 2026
Model Releases

Also out today: You can now directly configure the effort level and adaptive thinking in Code (/effort) and Cowork! Effort allows you to tun…

DGX agent

Also out today: You can now directly configure the effort level and adaptive thinking in Code (/effort) and Cowork! Effort allows you to tune Claude's intelligence vs token spend, trading off capabili

model-releasesboris-cherny--x
28 May 2026
Model Releases

An Enhanced Large Neighborhood Search Approach for the Capacitated Facility Location Problem with Incompatible Customers

DGX agent

arXiv:2605.28337v1 Announce Type: new Abstract: A new variant of the classic capacitated facility location problem, which considers incompatibilities between customers, has recently been introduced in

model-releasesarxiv-cs-ai
28 May 2026
Model Releases

Analyzing Quality-Latency-Resource Trade-offs in a Technical Documentation RAG Assistant Using LoRA Adaptation

DGX agent

arXiv:2605.28222v1 Announce Type: new Abstract: We study quality-latency-resource trade-offs in a documentation-grounded retrieval-augmented generation (RAG) system that uses Low-Rank Adaptation (LoRA

model-releasesarxiv-cs-cl
28 May 2026
← Previous
1…250251252253254…476
Next →