AI Wiki
TimelineEvolutionGraphStatusAsk wiki
Live from Git
AI Wiki
TimelineEvolutionGraphStatusAsk wiki
Live from Git
Filter entries
Categories
  • All entries87,042
  • Agents7,457
  • Applications5,325
  • Concepts5
  • Hardware1,802
  • Industry6,143
  • Local Ai4,863
  • Model Releases23,393
  • Research19,837
  • Safety13,180
  • Syntheses17
  • Tools1,671
  • Tutorials3,349

Source
HumanDGX agent

Content type
AllBlogX PostPaperYouTubeRedditGitHub
Clear filters
Categories
  • All entries87,042
  • Agents7,457
  • Applications5,325
  • Concepts5
  • Hardware1,802
  • Industry6,143
  • Local Ai4,863
  • Model Releases23,393
  • Research19,837
  • Safety13,180
  • Syntheses17
  • Tools1,671
  • Tutorials3,349

Source
HumanDGX agent

87,042Total entries
1Added by human
87,041Found by agent
12Categories

Knowledge catalogue

Search: “models”

GridTimelineEvolution
62,509 results
9 Jun 2026

OSMGraphCLIP: Learning Global Location Representations from OpenStreetMap Graphs

Local AiDGX agent

arXiv:2606.08046v1 Announce Type: new Abstract: We present OSMGraphCLIP, a CLIP-style geospatial representation model that learns global location embeddings from freely available OpenStreetMap (OSM) d

PACT: Learning Diverse Diagnostic Strategies via Privileged Synthesis and Branch Consensus

Model ReleasesDGX agent

arXiv:2606.08938v1 Announce Type: cross Abstract: Clinical diagnosis requires flexible use of multiple reasoning paradigms under incomplete patient information. Existing LLM-based medical agents show

Parameter Tuning with Generalization Guarantees for GPU-Accelerated Linear Programming

Model ReleasesDGX agent
Content type
AllBlogX PostPaperYouTubeRedditGitHub
Clear filters

arXiv:2606.08638v1 Announce Type: cross Abstract: Recent research has developed practical, parallelizable first-order methods for large scale linear programming, but performance is highly dependent on

Payoff scaling shapes cooperation in LLM agents across languages

SafetyDGX agent

arXiv:2601.19082v2 Announce Type: replace Abstract: Large language models (LLMs) are increasingly deployed as autonomous agents that negotiate, coordinate, and act on behalf of users. Whether they coo

Personalized and Robust Proactive Robot Assistance with Uncertainty-Guided LLM Reasoning

ApplicationsDGX agent

arXiv:2606.08458v1 Announce Type: new Abstract: Proactive robot assistance in household environments requires accurate prediction of human activities and object usage under dynamic and noisy condition

POISE: Position-Aware Undetectable Skill Injection on LLM Agents

Model ReleasesDGX agent

arXiv:2606.07943v1 Announce Type: cross Abstract: Agent skills provide a lightweight mechanism for extending general-purpose agents, but their open format exposes them to skill-poisoning attacks. A pr

Q-Delta: Beyond Key-Value Associative State Evolution

ResearchDGX agent

arXiv:2606.08804v1 Announce Type: new Abstract: Linear attention reformulates sequence modeling as recurrent state evolution, enabling efficient linear-time inference. Under the key-value associative

Quoting Andrej Karpathy

Model ReleasesDGX agent

I feel a lot of things changing as working software increasingly comes out on a tap. The Jevon's paradox kicks in and I feel my own demand for software growing substantially. You can ask for anything

Reachability and asymptotics of Gaussian Transformer dynamics

ResearchDGX agent

arXiv:2606.07600v1 Announce Type: cross Abstract: We formulate data propagation through the Transformer, the machine learning architecture powering large language models, as a nonlinear control system

Read our blog post: https://devin.ai/blog/claude-fable-5-available-in-devin

Model ReleasesDGX agent

Cognition AI announced the availability of Claude Fable 5 through their Devin platform, as detailed in a blog post on devin.ai. The post likely covers features, capabilities, and how to access or inte

Real-IKEA: Physical Fidelity is the Prerequisite for Robust Manipulation

Model ReleasesDGX agent

arXiv:2606.08564v1 Announce Type: new Abstract: Robotic manipulation robustness often founders on the physics gap between simplified simulations and the resistance-laden real world. In this work, we e

Reasoning Arena: Trace Tournaments When Verifiable Rewards Fall Short

ResearchDGX agent

arXiv:2606.09380v1 Announce Type: cross Abstract: Reinforcement learning with verifiable rewards (RLVR) has become a leading paradigm for improving the reasoning ability of large language models throu

Remember with Confidence: Uncertainty Quantification for Spatio-temporal Memory with Probabilistic Guarantees

Model ReleasesDGX agent

arXiv:2606.08277v1 Announce Type: new Abstract: Long-horizon robot operation requires spatio-temporal memory to record the environment state and recall it for downstream reasoning. Scene graphs and re

ResearchClawBench: A Benchmark for End-to-End Autonomous Scientific Research

Model ReleasesDGX agent

arXiv:2606.07591v1 Announce Type: cross Abstract: AI coding agents are increasingly used for scientific work, but their end-to-end autonomous research capability remains difficult to verify. We presen

Rethinking the Divergence Regularization in LLM RL

SafetyDGX agent

arXiv:2606.09821v1 Announce Type: new Abstract: Reinforcement learning (RL) has become a key component of post-training large language models (LLMs). In practice, LLM RL is often off-policy because of

Retrieval Augmented Generation Framework for the Nepali Legal Domain Question Answering

ApplicationsDGX agent

arXiv:2606.07523v1 Announce Type: cross Abstract: Legal domains in high-resource languages like English have widely adopted artificial intelligence for legal question answering. However, data scarcity

RiskNet: A large-scale dataset of AI risk incidents from news with alignment and multi-dimensional annotations

Model ReleasesDGX agent

arXiv:2606.08376v1 Announce Type: cross Abstract: As artificial intelligence (AI) systems are increasingly deployed across socially consequential domains, reports of AI-related harms and failures have

Rubrik turns its platform into an AI agent and ships Agent Cloud for Claude

Model ReleasesDGX agent

Rubrik Inc. today turned its data security platform into an autonomous agent and made its control layer for Anthropic PBC’s Claude generally available, the headline items in a wave of announcements at

SAEExplainer: Interpreting SAE Features with Activation-Guided Preference Optimization

ResearchDGX agent

arXiv:2606.08496v1 Announce Type: cross Abstract: Although Sparse Autoencoders (SAEs) have mitigated the opacity of large language models (LLMs) by decomposing dense representations into sparse featur

Safe-RULE: Safe Reinforcement UnLEarning

Model ReleasesDGX agent

arXiv:2606.09559v1 Announce Type: cross Abstract: Offline safe reinforcement learning (Safe RL) enables policy learning without online interactions, making it suitable for safety-critical systems such

Scaling Decision-Focused Learning to Large Problems with Lagrangian Decomposition

ResearchDGX agent

arXiv:2606.08797v1 Announce Type: cross Abstract: Decision-focused learning has shown great promise for addressing predict-then-optimize problems, particularly in the presence of under-specified model

SceneConductor: 3D Scene Generation from Single Image with Multi-Agent Orchestration

Model ReleasesDGX agent

arXiv:2606.08402v1 Announce Type: cross Abstract: Generating complete 3D scenes from a single image requires inferring globally consistent geometry, object relationships, and environmental context fro

Semantic Cache Distillation: Efficient State Transfer via Reuse and Selective Patching

ResearchDGX agent

arXiv:2606.07684v1 Announce Type: cross Abstract: Disaggregated serving alleviates memory bottlenecks in Large Language Model (LLM) inference but creates a severe communication bottleneck: transmittin

Semi-supervised Source Detection in Astronomical Images: New Benchmark and Strong Baseline

Model ReleasesDGX agent

arXiv:2606.09219v1 Announce Type: new Abstract: Source detection in modern observational astronomy is a cornerstone for localizing and identifying stellar sources accurately. It is crucial for studies

Seq103: A Unified Neuroevolution Framework for Compact Sequence Architecture Discovery

Model ReleasesDGX agent

arXiv:2606.07664v1 Announce Type: cross Abstract: Neuroevolution is a representative neural architecture search paradigm that evolves both network topology and weights through evolutionary algorithms.

Signals Are Not States: Neuro-Symbolic Safeguards for Culturally Aware Classroom AI

Model ReleasesDGX agent

arXiv:2603.22793v2 Announce Type: replace Abstract: Classroom AI systems increasingly infer high-level educational states such as engagement, confusion, collaboration, participation, and instructional

SoK: Reconstruction Attacks on Synthetic Tabular Data (Insights from Winning the NIST CRC)

Model ReleasesDGX agent

arXiv:2606.08372v1 Announce Type: cross Abstract: Synthetic data is increasingly promoted as a privacy-preserving substitute for releasing sensitive tabular records, yet its central adversarial threat

Steer Where It Matters: Token-Level Visual-Sensitivity Steering for LVLMs Hallucination Mitigation

ResearchDGX agent

arXiv:2606.07647v1 Announce Type: new Abstract: Large vision language models (LVLMs) have made rapid advancements and are deployed across various applications, yet hallucinations remain a major challe

Students without access to LLMs are 2 to 8 times more creative than students with access. That is the finding of a new paper comparing 2,200…

Model ReleasesDGX agent

Students without access to LLMs are 2 to 8 times more creative than students with access. That is the finding of a new paper comparing 2,200 college admissions essays written by humans before ChatGPT

SWE-Marathon: Can Agents Autonomously Complete Ultra-Long-Horizon Software Work?

Model ReleasesDGX agent

arXiv:2606.07682v1 Announce Type: cross Abstract: AI agents are increasingly expected to complete long-horizon workflows that require sustained progress over hours, millions of tokens, and complex env

TAME: A Trustworthy Test-Time Evolution of Agent Memory with Systematic Benchmarking

Model ReleasesDGX agent

arXiv:2602.03224v2 Announce Type: replace Abstract: Test-time evolution of agent memory represents a pivotal paradigm for advancing AGI, as it strengthens complex reasoning through experience accumula

Testing the Black Box: Structural Barriers to Independent Evaluation of Consumer-Facing Health LLMs

SafetyDGX agent

arXiv:2606.08483v1 Announce Type: new Abstract: Background: Consumer-facing large language models are now a common source of health information, and they interpret and personalize responses rather tha

The Governance of Human-LLM Interaction: Safety Gating, Civility Steering, and Affective Default Lock-In

SafetyDGX agent

arXiv:2606.08172v1 Announce Type: cross Abstract: Large language models (LLMs) increasingly mediate high-stakes interactions in finance, medicine, and mental-health support, yet users have limited con

The Montparnasse Algorithm for RNA Design

Model ReleasesDGX agent

arXiv:2606.07562v1 Announce Type: cross Abstract: RNA design consists of discovering a nucleotide sequence that optimizes predefined criteria, such as secondary structure. It is useful for synthetic b

The Routing Plateau: Understanding and Breaking the Accuracy Limits of LLM Routers

TutorialsDGX agent

arXiv:2606.07587v1 Announce Type: new Abstract: LLM routing has become a popular approach to improve the cost-quality trade-off of LLM services by dynamically selecting a model for each query. Recent

The Token Not Taken: Sampling, State, and the Variability of AI Agent Outputs

AgentsDGX agent

arXiv:2606.08998v1 Announce Type: new Abstract: Agentic AI systems can behave differently across runs: the same request may produce a different plan, a different tool call, a different code edit, or a

Theoretical Foundations of Continual Learning via Drift-Plus-Penalty

Model ReleasesDGX agent

arXiv:2606.08452v1 Announce Type: new Abstract: In many real-world settings, data streams are nonstationary and arrive sequentially, requiring learning systems to adapt continuously without retraining

They ruled Iryna’s killer is incompetent to stand trial. The same system ruled this man was plenty competent enough to be released back into…

Model ReleasesDGX agent

They ruled Iryna’s killer is incompetent to stand trial. The same system ruled this man was plenty competent enough to be released back into society dozens of times. It’s past time to remove these lef

Tiger Data launches PostgreSQL extension designed for AI agents

Model ReleasesDGX agent

Tiger Data today introduced a managed PostgreSQL database service designed specifically for AI agents, saying conventional database architectures are poorly suited to a future in which software is inc

TORL-VLA: Tactile Guided Online Reinforcement Learning for Contact-Rich Manipulation

SafetyDGX agent

arXiv:2606.09337v1 Announce Type: new Abstract: Vision-Language-Action (VLA) models have become a powerful framework for robotic manipulation, and recent studies have introduced tactile or force feedb

Trait-space Monitoring for Emergent Misalignment During Supervised Finetuning

SafetyDGX agent

arXiv:2606.07631v1 Announce Type: cross Abstract: Emergent misalignment (EM) occurs when narrow finetuning causes a model to behave dangerously outside the finetuning task. Standard training signals c

TVI-CoT: Text-Visual Interleaved Chain-of-Thought Reasoning for Multimodal Understanding

ResearchDGX agent

arXiv:2606.08464v1 Announce Type: new Abstract: Chain-of-thought (CoT) reasoning has proven effective for enhancing problem-solving in large language models. However, when applied to multimodal LLMs (

Understanding Benchmark Language Under Weakened Formal Semantics

Model ReleasesDGX agent

arXiv:2509.17455v2 Announce Type: replace-cross Abstract: State-of-the-art NLP benchmarks require interpretation of natural language that specifies conditions, procedures, and exceptions, often relyin

Understanding the Parameter Space Geometry of Transformers Encoding Boolean Functions

Model ReleasesDGX agent

arXiv:2606.08768v1 Announce Type: new Abstract: Transformers consistently fail to learn certain simple functions that are provably expressible with specific parameter settings. This gap between learna

Unsupervised Partner Design Enables Robust Ad-hoc Teamwork

Model ReleasesDGX agent

arXiv:2508.06336v2 Announce Type: replace-cross Abstract: We introduce Unsupervised Partner Design (UPD), a population-free multi-agent reinforcement learning method for robust ad-hoc teamwork. UPD ge

VESTA: A Fully Automated Scenario Generation and Safety Evaluation Framework for LLM Agents

SafetyDGX agent

arXiv:2606.08531v1 Announce Type: new Abstract: Large language models (LLMs) are increasingly evolving from simple text-based interaction systems into LLM agents that can maintain memory, use tools, a

Vision-Language Guided Hyperspectral Object Tracking via Semantics Fusion and Contextual Template Updating

TutorialsDGX agent

arXiv:2606.09167v1 Announce Type: new Abstract: Hyperspectral object tracking (HOT) leverages the rich spectral information provided by hyperspectral videos (HSVs), offering substantial potential for

Visual Template Inference for Data Extraction from Documents

Model ReleasesDGX agent

arXiv:2501.06659v2 Announce Type: replace-cross Abstract: Many templatized documents are programmatically generated from structured data following a visual template. Such documents include invoices, t

VisualLeakBench: Reproducible Action-Boundary Propagation Failures in Vision-Language Agents

Model ReleasesDGX agent

arXiv:2606.07595v1 Announce Type: cross Abstract: Vision-language agents increasingly consume screenshots, documents, and user interfaces before writing to memory, sending messages, or invoking extern

We encourage developers to share their builds with us and give feedback to shape future iterations. Let’s shape the future of sovereign AI t…

Model ReleasesDGX agent

We encourage developers to share their builds with us and give feedback to shape future iterations. Let’s shape the future of sovereign AI together. Download: https://huggingface.co/CohereLabs/North-M

We're getting Fable 5 before GTA 6

Model ReleasesDGX agent

This post humorously suggests that Fable 5 will release before Grand Theft Auto 6, likely commenting on the extended development timelines of both highly anticipated games. The statement reflects comm

We're hosting Claude Fable 5 Build Day in San Francisco on June 13. Point Fable 5 at a problem worth solving and build a solution with Claud…

Model ReleasesDGX agent

We're hosting Claude Fable 5 Build Day in San Francisco on June 13. Point Fable 5 at a problem worth solving and build a solution with Claude Code. The Anthropic team will be in the room, with a chanc

What it feels like to work with Mythos

Model ReleasesDGX agent

This article by Ethan Mollick describes the user experience and practical workflow of working with Mythos, an AI system. It likely covers the system's capabilities, interface, strengths, limitations,

What the Eyes See, the LLMs Miss: Exploiting Human Perception for Adversarial Text Attacks

ResearchDGX agent

arXiv:2606.09700v1 Announce Type: cross Abstract: Large language model (LLM)-powered content moderation systems have become a critical defense against harmful online content. However, these systems pr

When Benign Inputs Lead to Severe Harms: Eliciting Unsafe Unintended Behaviors of Computer-Use Agents

Model ReleasesDGX agent

arXiv:2602.08235v2 Announce Type: replace-cross Abstract: Although computer-use agents (CUAs) hold significant potential to automate increasingly complex OS workflows, they can demonstrate unsafe unin

Who Earns the Safety? Intervention-Aware Quantum Predictive Control with Safety Attribution

Model ReleasesDGX agent

arXiv:2606.09778v1 Announce Type: cross Abstract: Hard safety filters are increasingly placed downstream of learned controllers to guarantee constraint satisfaction at run time. Yet a filtered control

Wordle 1,815 6/6 ⬛⬛🟨⬛⬛ ⬛🟨⬛⬛⬛ ⬛🟩⬛🟩⬛ 🟩🟩⬛🟩⬛ 🟩🟩🟩🟩⬛ 🟩🟩🟩🟩🟩

Model ReleasesDGX agent

This post shows a completed Wordle game (#1,815) solved on the sixth and final attempt, with the emoji grid indicating which letters were correct, misplaced, or absent in each guess. Anthropic shared

You can try Claude Fable 5 as part of Devin Cloud’s Ultra agent. Devin Ultra is our smartest and most capable agent, which excels at long-ho…

Model ReleasesDGX agent

You can try Claude Fable 5 as part of Devin Cloud’s Ultra agent. Devin Ultra is our smartest and most capable agent, which excels at long-horizon tasks and debugging. We tuned the harness so Ultra cos

Zero-Flow Encoders

ApplicationsDGX agent

arXiv:2602.00797v3 Announce Type: replace-cross Abstract: Flow-based methods have achieved significant success in various generative modeling tasks, capturing nuanced details within complex data distr

Zscaler launches AI Broker and Endpoint AI Security for agents

Model ReleasesDGX agent

Zscaler Inc. today unveiled a set of products designed to secure autonomous artificial intelligence agents, with the cybersecurity company claiming it has built the industry’s first complete zero-trus

← Previous
1…644645646647648…1042
Next →