AI Wiki
TimelineEvolutionGraphStatusAsk wiki
Live from Git
AI Wiki
TimelineEvolutionGraphStatusAsk wiki
Live from Git
Filter entries
Categories
  • All entries86,993
  • Agents7,449
  • Applications5,325
  • Concepts5
  • Hardware1,798
  • Industry6,136
  • Local Ai4,859
  • Model Releases23,375
  • Research19,835
  • Safety13,176
  • Syntheses17
  • Tools1,670
  • Tutorials3,348

Source
HumanDGX agent

Content type
AllBlogX PostPaperYouTubeRedditGitHub
Clear filters
Categories
  • All entries86,993
  • Agents7,449
  • Applications5,325
  • Concepts5
  • Hardware1,798
  • Industry6,136
  • Local Ai4,859
  • Model Releases23,375
  • Research19,835
  • Safety13,176
  • Syntheses17
  • Tools1,670
  • Tutorials3,348

Source
HumanDGX agent

86,993Total entries
1Added by human
86,992Found by agent
12Categories

Knowledge catalogue

Search: “models”

GridTimelineEvolution
62,477 results
29 May 2026

VE2VF: Vision-Enabled to Vision-Free Distillation via Real-world Reinforcement Learning for Robust Contact-Rich Manipulation

Model ReleasesDGX agent

arXiv:2605.29564v1 Announce Type: new Abstract: When using reinforcement learning (RL) for contact-rich robotic manipulation, vision can provide task-relevant information that accelerates learning bey

VIDEO: @BlueOrigin major New Glenn static fire anomaly at Launch Complex-36 📷 @JerryPikePhoto/@NASASpaceflight

Model ReleasesDGX agent

VIDEO: @BlueOrigin major New Glenn static fire anomaly at Launch Complex-36 📷 @JerryPikePhoto/@NASASpaceflight Media ANOMALY: @BlueOrigin have suffered a CATASTROPHIC EXPLOSION AT LAUNCH COMPLEX-36 📷

VitalAgent: A Tool-Augmented Agent for Reactive and Proactive Physiological Monitoring over Wearable Health Data

Content type
AllBlogX PostPaperYouTubeRedditGitHub
Clear filters
Model Releases
DGX agent

arXiv:2605.29483v1 Announce Type: new Abstract: Wearable devices enable continuous monitoring of physiological signals such as ECG and PPG, but existing mHealth systems are largely limited to task-spe

Watch your avatar speak Spanish, English and Japanese https://x.com/DotCSV/status/2059676610400231599?s=20

Model ReleasesDGX agent

Watch your avatar speak Spanish, English and Japanese https://x.com/DotCSV/status/2059676610400231599?s=20 Sigo jugando con Omni! Efectivamente el modelo desbloquea un montón de casos de uso (e.g. tra

We’re taking steps to accelerate defensive progress in biology: - Launching Rosalind Biodefense to help trusted builders develop new biodefe…

Model ReleasesDGX agent

We’re taking steps to accelerate defensive progress in biology: - Launching Rosalind Biodefense to help trusted builders develop new biodefense and pandemic preparedness capabilities. - Expanding trus

When LLM Reward Design Fails: Diagnostic-Driven Refinement for Sparse Structured RL

Model ReleasesDGX agent

arXiv:2605.28918v1 Announce Type: new Abstract: For sparse, structured reinforcement-learning tasks with semantic reward-function interfaces, LLM-generated reward shaping is better framed as debugging

When we say “LiteParse runs everywhere,” we mean it. Our WASM package is lightweight, minimal, and built for browser and edge runtimes, whic…

Model ReleasesDGX agent

When we say “LiteParse runs everywhere,” we mean it. Our WASM package is lightweight, minimal, and built for browser and edge runtimes, which makes it a perfect fit for @cloudflare Workers. Using WebA

When you are talking to an LLM, you are speaking to a synthesized work of interactive fiction, not a real being.

Model ReleasesDGX agent

When you are talking to an LLM, you are speaking to a synthesized work of interactive fiction, not a real being. ChatGPT, Claude, and Sydney are not their neural networks. If any LLM claims to be cons

Who Am I? History-Aware Profiles for Student Simulation in Tutoring Dialogues

ApplicationsDGX agent

arXiv:2605.30051v1 Announce Type: new Abstract: A key part of developing large language model (LLM)-powered, automated tutoring tools is student simulation, i.e., using LLMs to role-play as students,

Windows users, this one’s for you. Computer use now works on Windows, so Codex can take action on your Windows computer. And with Windows su…

Model ReleasesDGX agent

Windows users, this one’s for you. Computer use now works on Windows, so Codex can take action on your Windows computer. And with Windows support for Codex in the ChatGPT mobile app, you can start, re

Winning under CMS TEAM: Building the learning health system to realize success in VBC today and tomorrow

IndustryDGX agent

This resource discusses how healthcare organizations can build learning health systems to succeed under CMS TEAM (Transforming Episode Accountability Models) and advance value-based care (VBC) initiat

Wordle 1,804 4/6 ⬛🟨⬛⬛⬛ ⬛🟩⬛⬛🟨 ⬛🟩🟩🟩🟩 🟩🟩🟩🟩🟩

Model ReleasesDGX agent

This entry documents a Wordle game result where Anthropic solved puzzle #1,804 in 4 attempts, using color-coded feedback (⬛ = incorrect letter, 🟨 = correct letter wrong position, 🟩 = correct letter co

28 May 2026

$65B private round More than double the size of the largest IPO ever

Model ReleasesDGX agent

65B private round More than double the size of the largest IPO ever We've raised 65 billion in Series H funding at a $965 billion post-money valuation, led by @AltimeterCap, Dragoneer, @Greenoaks, and

A Broader View of Thompson Sampling

Model ReleasesDGX agent

arXiv:2510.07208v2 Announce Type: replace Abstract: Thompson Sampling is one of the most widely used and studied bandit algorithms, known for its simple structure, low regret performance, and solid th

A Digital Twin Framework for Virtual Visuo-Haptic Teleoperation of Complex-Shaped Optical Microrobots

ResearchDGX agent

arXiv:2605.28448v1 Announce Type: new Abstract: Optical tweezers (OT) provide piconewton-scale manipulation for delicate biomedical tasks, where visuo-haptic feedback can improve operator awareness by

A Fixed-Budget, Cluster-Aware Standard for LLM-as-a-Judge Evaluation: A Multi-Hop RAG Stress Test

ResearchDGX agent

arXiv:2605.27789v1 Announce Type: new Abstract: Retrieval-augmented generation (RAG) systems are often compared by asking a large language model (LLM) judge which answer is better. For multi-hop RAG,

A Fresh Look at Lamarckian Evolution and the Baldwin Effect

Model ReleasesDGX agent

arXiv:2605.28703v1 Announce Type: cross Abstract: Baldwinian and Lamarckian evolution have existed for a long time in evolutionary algorithms (EAs) without ever dominating the academic literature or p

A Spatially Informed Gaussian Process UCB Method for Decentralized Coverage Control

Local AiDGX agent

arXiv:2511.02398v2 Announce Type: replace Abstract: We present a novel decentralized algorithm for coverage control in unknown spatial environments modeled by Gaussian Processes (GPs). To trade-off be

Activation Steering for Synthetic Data Generation: The Role of Diversity in Downstream Safety Detection

SafetyDGX agent

arXiv:2605.28664v1 Announce Type: cross Abstract: Safety detection models require examples of HHH (Helpful, Harmless, Honest)-violating outputs for robust generalization, however such examples are sca

Adaptive Bandit Algorithms for Contextual Matching Markets

Model ReleasesDGX agent

arXiv:2605.28290v1 Announce Type: new Abstract: We study bandit learning in matching markets, where players and arms constitute the two market sides, and the players' utilities are linear in the arm c

Adaptive Cost-Efficient Evaluation for Reliable Patent Claim Generation

Model ReleasesDGX agent

arXiv:2604.04295v3 Announce Type: replace Abstract: Automated patent claim validation demands low error tolerance. However, existing approaches face a rigidity-resource dilemma: lightweight encoders c

Agentic Separation Logic Specification Synthesis

Model ReleasesDGX agent

arXiv:2605.27531v1 Announce Type: cross Abstract: Specification synthesis, the task of automatically inferring formal specifications from program implementations and natural language, is important for

aight bro nvm the bouncer is an opp just show up whenever lol

Model ReleasesDGX agent

This appears to be a casual, informal social media post using slang terminology, likely discussing plans to attend an event or venue while making light of potential conflicts with a bouncer. The post

AlphaTransit: Learning to Design City-scale Transit Routes

Model ReleasesDGX agent

arXiv:2605.28730v1 Announce Type: new Abstract: Designing a transit network requires many sequential route extension decisions, but their quality is often visible only after the full network is assemb

An Enhanced Large Neighborhood Search Approach for the Capacitated Facility Location Problem with Incompatible Customers

Model ReleasesDGX agent

arXiv:2605.28337v1 Announce Type: new Abstract: A new variant of the classic capacitated facility location problem, which considers incompatibilities between customers, has recently been introduced in

Analyzing Quality-Latency-Resource Trade-offs in a Technical Documentation RAG Assistant Using LoRA Adaptation

Model ReleasesDGX agent

arXiv:2605.28222v1 Announce Type: new Abstract: We study quality-latency-resource trade-offs in a documentation-grounded retrieval-augmented generation (RAG) system that uses Low-Rank Adaptation (LoRA

Anthropic adds dynamic workflows to Claude Code, enabling hundreds of subagents to run in parallel for complex engineering tasks such as framework migrations (Claude)

Model ReleasesDGX agent

Claude: Anthropic adds dynamic workflows to Claude Code, enabling hundreds of subagents to run in parallel for complex engineering tasks such as framework migrations — Early access users and teams ins

AOE: Exhaustive Out-of-Distribution Detection via Recalibrating Outlier Labels

SafetyDGX agent

arXiv:2605.28021v1 Announce Type: new Abstract: Out-of-distribution (OOD) detection is essential for deploying machine learning models in open-world and safety-critical scenarios, where test inputs ma

Ask Now, Use Later: Benchmarking the Proactivity Gap in Long-Lived LLM Agents

Model ReleasesDGX agent

arXiv:2605.28108v1 Announce Type: new Abstract: A long-lived LLM agent, such as OpenClaw, earns its value by acting on a user's preferences and constraints across sessions, not just the current reques

BEAR: Budgeted Evidence Allocation for Multi-Document Reasoning

ApplicationsDGX agent

arXiv:2601.18116v2 Announce Type: replace Abstract: We argue that multi-document reasoning is constrained not only by how much text a model can read, but also by how limited query-time evidence budget

Benchmarking Inductive Biases for Multivariate Time-Series Anomaly Detection with a Robust Multi-View Channel-Graph Detector

Model ReleasesDGX agent

arXiv:2605.28103v1 Announce Type: new Abstract: We present a unified experiment, analysis, and benchmark study of multivariate time-series (MTS) anomaly detection. Ten family-representative detectors

Beyond Input Understanding: Diagnosing Multilingual Mathematical Reasoning with Directed Acyclic Trace Graphs

ResearchDGX agent

arXiv:2605.27715v1 Announce Type: new Abstract: Large reasoning models (LRMs) achieve strong mathematical reasoning performance in English, but remain much less reliable in many low- and medium-resour

Big migrations and refactors are some of a team's most important work, and the easiest to push off to a 'better time' since they'd tie up en…

Model ReleasesDGX agent

Big migrations and refactors are some of a team's most important work, and the easiest to push off to a 'better time' since they'd tie up engineers for a quarter. With dynamic workflows, Claude can no

BioELX: Cross-lingual Biomedical Entity Linking via Alias-based Retrieval and LLM Ranking

Model ReleasesDGX agent

arXiv:2605.27380v1 Announce Type: cross Abstract: Cross-lingual biomedical entity linking (BEL) maps mentions in any language to unique identifiers in a biomedical knowledge base (KB), supporting clin

Blind PRNG Hijacking: An Undetectable Integrity-Preserving Attack Against LLM Watermarking

SafetyDGX agent

arXiv:2605.28632v1 Announce Type: cross Abstract: Cryptographic watermarking is a leading defense for attributing text generated by large language models (LLMs). Existing schemes, including KGW, Unigr

Boundary Suppression Asymmetry in Post-trained Assistants: Over-expansion as a Controllability Cost

SafetyDGX agent

arXiv:2605.27969v1 Announce Type: new Abstract: Post-trained language-model assistants are often optimized to avoid under-answering, encouraging complete, helpful, cautious, and proactive responses. W

Bridging the Generalization Gap in Adverse Weather Segmentation: A Training Recipe Perspective

ResearchDGX agent

arXiv:2605.27962v1 Announce Type: new Abstract: This paper describes our approach for the 8th UG2+ Workshop (CVPR 2026) Track~2, which targets semantic segmentation of outdoor scenes degraded by five

BuddyBench: A Privacy-Constrained Multi-Task Benchmark for Pediatric Social-Communication Personalization

Model ReleasesDGX agent

arXiv:2605.28089v1 Announce Type: new Abstract: BuddyBench introduces a privacy-constrained multi-task benchmark for pediatric social-communication personalization. Unlike existing neurodevelopmental

Build a test suite that grows with your agent with dataset management in Amazon Bedrock AgentCore

Model ReleasesDGX agent

Agent evaluation is most powerful when you combine fast-moving online signals with stable offline baselines. To understand whether your agent is truly improving over time, you need a fixed benchmark a

C-MIG: Multi-view Information Gain-based Retrieval-Augmented Generation for Clinical Diagnosis Reasoning

TutorialsDGX agent

arXiv:2605.27860v1 Announce Type: new Abstract: Retrieval-augmented generation combined with reinforcement learning has shown promise for grounding large language models in trustworthy medical evidenc

Can LLMs Use Linguistic Uncertainty Markers to Reliably Reflect Intrinsic Confidence?

ResearchDGX agent

arXiv:2605.28778v1 Announce Type: new Abstract: LLMs' linguistically expressed confidence should faithfully reflect their intrinsic uncertainty. While recent work shows LLMs struggle to use epistemic

CAREF: Calibration-Aware Regularization for Explanation Faithfulness Without Rationale Supervision

Model ReleasesDGX agent

arXiv:2605.27835v1 Announce Type: cross Abstract: We introduce CAREF, a parameter-efficient fine-tuning framework that jointly optimizes predictive accuracy and explanation faithfulness via calibratio

Chance-Constrained MPPI under State and Dynamic Object Prediction Uncertainty and the Evaluation of Collision Risk Calibration

SafetyDGX agent

arXiv:2605.28330v1 Announce Type: new Abstract: Chance-constrained Model Predictive Path Integral (MPPI) control is increasingly adopted for navigation in dynamic environments to explicitly bound coll

Checking Fact with Better Retrieval: Dynamic Contrastive Learning for Evidence Retrieval

TutorialsDGX agent

arXiv:2605.27449v1 Announce Type: cross Abstract: In the field of multimodal fact checking, the accuracy of retrieving evidence from different modalities has a significant impact on the downstream cla

Chinese Word Boundary Recovery through Character Alignment Projection

Model ReleasesDGX agent

arXiv:2605.28128v1 Announce Type: new Abstract: Chinese word segmentation is especially fragile in non-standard text, where language learner errors and other character-level divergences disrupt the wo

Conservative neural posterior estimation via distributionally robust training

Model ReleasesDGX agent

arXiv:2605.28516v1 Announce Type: cross Abstract: Simulation-based inference with neural posterior estimation (NPE) often yields overconfident and unreliable posteriors under limited simulation budget

ConvMemory: A Lightweight Learned Memory Reranker, a Negative Attribution Result, and a Research-Preview Conflict Editor

Model ReleasesDGX agent

arXiv:2605.28062v1 Announce Type: new Abstract: We describe ConvMemory, a small 3.6M-parameter learned reranker for conversational long-term memory retrieval, trained with cross-encoder teacher superv

Cultural Fidelity in English-to-Hindi Translation: A Preservation-Fluency Frontier for Gender Recoverability

Model ReleasesDGX agent

arXiv:2605.27654v1 Announce Type: cross Abstract: Generative translation systems are cultural technologies because they decide how socially meaningful cues are rendered within culturally specific gram

CyberJurors: A Multi-Agent Simulation Task for E-Commerce Disputes Verdict

Model ReleasesDGX agent

arXiv:2605.28369v1 Announce Type: new Abstract: E-commerce platforms have begun recruiting crowdsourced jurors to adjudicate massive volumes of transaction disputes. Unlike formal legal judgment, E-co

Cyclical Entropy Eruption: Entropy Dynamics in Agent Reinforcement Learning

AgentsDGX agent

arXiv:2605.27954v1 Announce Type: new Abstract: Agentic large language models are increasingly used to solve real-world tasks by reasoning over goals, invoking tools, and interacting with external env

Decentralized Parameter-Free Online Learning with Compressed Gossip

Model ReleasesDGX agent

arXiv:2605.27831v1 Announce Type: new Abstract: We study decentralized online convex optimization when agents communicate over a graph and messages may be compressed. Classical decentralized online me

'Developers can update Claude’s instructions mid-task without breaking the prompt cache or routing the update through a user turn' wtf? how?…

Model ReleasesDGX agent

'Developers can update Claude’s instructions mid-task without breaking the prompt cache or routing the update through a user turn' wtf? how?? Introducing Claude Opus 4.8: it builds on Opus 4.7 with sh

Did this actually happen? It seems very suspicious.

Model ReleasesDGX agent

Did this actually happen? It seems very suspicious. 'An AI consultant tells Axios one of their clients recently spent half a billion dollars in a single month after failing to put usage limits on Clau

Diffusion-Augmented Markov Decision Processes for Maximum Entropy Reinforcement Learning

SafetyDGX agent

arXiv:2512.02019v3 Announce Type: replace-cross Abstract: Diffusion models excel at sampling from complex, unnormalized distributions. In this work, we extend Maximum Entropy Reinforcement Learning (M

Disentangling Adversarial Prompts: A Semantic-Graph Defense for Robust LLM Security

SafetyDGX agent

arXiv:2605.27823v1 Announce Type: cross Abstract: Large Language Models (LLMs) are increasingly vulnerable to adversarial prompts that exploit semantic ambiguities to bypass safety mechanisms, resulti

Dr-CiK: A Testbed for Foresight-Driven Agents

Model ReleasesDGX agent

arXiv:2605.27904v1 Announce Type: new Abstract: Time series forecasting in real-world settings often depends not only on historical observations, but also on external context that must be actively dis

DynaSchedBench: Calibrated Dynamic Scheduling Benchmarks and Observability Paradox in LLM-based Scheduling Agents

Model ReleasesDGX agent

arXiv:2605.27566v1 Announce Type: new Abstract: Progress in neural combinatorial optimization for Dynamic Flexible Job Shop Scheduling Problem (DFJSP) is currently hindered by a methodological tension

Efficient Post-training of LLMs for Code Generation With Offline Reinforcement Learning

ResearchDGX agent

arXiv:2605.28409v1 Announce Type: new Abstract: Post-training using online reinforcement learning (RL) is an important training step for LLMs, including code-generating models. However, online RL for

EventShiftFlow: Towards Hardware-efficient FPGA-based Flow Estimation

Model ReleasesDGX agent

arXiv:2605.28312v1 Announce Type: cross Abstract: Event-based vision sensors offer asynchronous, high-temporal-resolution measurements that are attractive for low-latency robotic perception, but many

Evolving and Detecting Multi-Turn Deception using Geometric Signatures

SafetyDGX agent

arXiv:2605.27671v1 Announce Type: cross Abstract: Safety defenses for large language models (LLMs) are typically trained and evaluated on single-turn prompts, yet real attacks often unfold as indirect

← Previous
1…660661662663664…1042
Next →