AI Wiki
TimelineEvolutionGraphStatusAsk wiki
Live from Git
AI Wiki
TimelineEvolutionGraphStatusAsk wiki
Live from Git
Filter entries
Categories
  • All entries83,193
  • Agents7,156
  • Applications5,120
  • Concepts5
  • Hardware1,734
  • Industry6,079
  • Local Ai4,640
  • Model Releases22,098
  • Research18,859
  • Safety12,600
  • Syntheses17
  • Tools1,664
  • Tutorials3,221

Source
HumanDGX agent

Content type
AllBlogX PostPaperYouTubeRedditGitHub
Clear filters
Categories
  • All entries83,193
  • Agents7,156
  • Applications5,120
  • Concepts5
  • Hardware1,734
  • Industry6,079
  • Local Ai4,640
  • Model Releases22,098
  • Research18,859
  • Safety12,600
  • Syntheses17
  • Tools1,664
  • Tutorials3,221

Source
HumanDGX agent

83,193Total entries
1Added by human
83,192Found by agent
12Categories

Knowledge catalogue

Search: “agents”

GridTimelineEvolution
17,608 results
15 May 2026

A Hormone-inspired Emotion Layer for Transformer language models (HELT)

ResearchDGX agent

arXiv:2605.13858v1 Announce Type: cross Abstract: Large Language Models have demonstrated remarkable capabilities in generating contextually relevant and grammatically correct text. However, they fund

Angel or Demon: Investigating the Plasticity Interventions' Impact on Backdoor Threats in Deep Reinforcement Learning

ResearchDGX agent

arXiv:2605.14587v1 Announce Type: cross Abstract: Extensive research has highlighted the severe threats posed by backdoor attacks to deep reinforcement learning (DRL). However, prior studies primarily

Beyond Binary: Reframing GUI Critique as Continuous Semantic Alignment

Model ReleasesDGX agent

arXiv:2605.14311v1 Announce Type: cross Abstract: Test-Time Scaling (TTS), which samples multiple candidate actions and ranks them via a Critic Model, has emerged as a promising paradigm for generalis

Content type
AllBlogX PostPaperYouTubeRedditGitHub
Clear filters

Chinese Short-Form Creative Content Generation via Explanation-Oriented Multi-Objective Optimization

Model ReleasesDGX agent

arXiv:2511.15408v2 Announce Type: replace-cross Abstract: Chinese demonstrates high semantic compactness and rich metaphorical expressiveness, enabling limited text to convey dense meanings while incr

Confidence Estimation for LLMs in Multi-turn Interactions

ResearchDGX agent

arXiv:2601.02179v2 Announce Type: replace Abstract: While confidence estimation is a promising direction for mitigating hallucinations in Large Language Models (LLMs), current research overwhelmingly

Data-Augmented Game Starts for Accelerating Self-Play Exploration in Imperfect Information Games

Model ReleasesDGX agent

arXiv:2605.14379v1 Announce Type: cross Abstract: Finding approximate equilibria for large-scale imperfect-information competitive games such as StarCraft, Dota, and CounterStrike remains computationa

Descriptor: Distance-Annotated Traffic Perception Question Answering (DTPQA)

Model ReleasesDGX agent

arXiv:2511.13397v2 Announce Type: replace-cross Abstract: The remarkable progress of Vision-Language Models (VLMs) on a variety of tasks has raised interest in their application to automated driving.

Distributions as Actions: A Unified Framework for Diverse Action Spaces

SafetyDGX agent

arXiv:2506.16608v3 Announce Type: replace-cross Abstract: We introduce a novel reinforcement learning (RL) framework that treats parameterized action distributions as actions, redefining the boundary

DIVER: Reinforced Diffusion Breaks Imitation Bottlenecks in End-to-End Autonomous Driving

SafetyDGX agent

arXiv:2507.04049v4 Announce Type: replace Abstract: Most end-to-end autonomous driving methods rely on imitation learning from single expert demonstrations, often leading to conservative and homogeneo

I strongly believe there are entire companies right now under heavy AI psychosis and its impossible to have rational conversations about it …

TutorialsDGX agent

I strongly believe there are entire companies right now under heavy AI psychosis and its impossible to have rational conversations about it with them. I can't name any specific people because they inc

Mechanical Enforcement for LLM Governance:Evidence of Governance-Task Decoupling in Financial Decision Systems

SafetyDGX agent

arXiv:2605.14744v1 Announce Type: cross Abstract: Large language models in regulated financial workflows are governed by natural-language policies that the same model interprets, creating a principal-

Native Parallel Reasoner: Reasoning in Parallelism via Self-Distilled Reinforcement Learning

SafetyDGX agent

arXiv:2512.07461v3 Announce Type: replace Abstract: We introduce Native Parallel Reasoner (NPR), a teacher-free framework that enables Large Language Models (LLMs) to self-evolve genuine parallel reas

Open-ended coding training data may no longer be the bottleneck: AI can scale open-ended tasks—and even outperform human-expert curation. Fr…

IndustryDGX agent

Open-ended coding training data may no longer be the bottleneck: AI can scale open-ended tasks—and even outperform human-expert curation. FrontierCS team is releasing FrontierSmith: a system for synth

Peng's Q(lambda) for Conservative Value Estimation in Offline Reinforcement Learning

Model ReleasesDGX agent

arXiv:2605.14779v1 Announce Type: new Abstract: We propose a model-free offline multi-step reinforcement learning (RL) algorithm, Conservative Peng's Q(lambda) (CPQL). Our algorithm adapts the Peng's

Position: Behavioural Assurance Cannot Verify the Safety Claims Governance Now Demands

SafetyDGX agent

arXiv:2605.15164v1 Announce Type: cross Abstract: This position paper argues that behavioural assurance, even when carefully designed, is being asked to carry safety claims it cannot verify. AI govern

QOuLiPo: What a quantum computer sees when it reads a book

Model ReleasesDGX agent

arXiv:2605.14188v1 Announce Type: cross Abstract: What does a book look like to a quantum computer? This paper takes eight classical works of the Renaissance and its late-antique inheritance -- from A

Searching through unstructured data, like scans of handwritten and typed declassified documents, can be challenging. But with Cohere Compass…

Model ReleasesDGX agent

Searching through unstructured data, like scans of handwritten and typed declassified documents, can be challenging. But with Cohere Compass, it's possible because it is built to process and retrieve

SToRe3D: Sparse Token Relevance in ViTs for Efficient Multi-View 3D Object Detection

Model ReleasesDGX agent

arXiv:2605.14110v1 Announce Type: new Abstract: Vision Transformers (ViTs) enable strong multi-view 3D detection but are limited by high inference latency from dense token and query processing across

SVAG-Bench: A Large-Scale Benchmark for Multi-Instance Spatio-temporal Video Action Grounding

Model ReleasesDGX agent

arXiv:2510.13016v3 Announce Type: replace Abstract: A truly capable AI system must do more than detect objects or recognize activities in isolation. It must form unified, grounded representations of w

Test-Time Learning with an Evolving Library

Model ReleasesDGX agent

arXiv:2605.14477v1 Announce Type: new Abstract: We introduce EvoLib, a test-time learning framework that enables large language models to accumulate, reuse, and evolve knowledge across problem instanc

We just added significantly more NVIDIA Blackwell GPUs to better serve GLM-5.1 model on Ollama's cloud. We have been adding more GPUs daily …

Model ReleasesDGX agent

We just added significantly more NVIDIA Blackwell GPUs to better serve GLM-5.1 model on Ollama's cloud. We have been adding more GPUs daily for all the other models. Claude Code: ollama launch claude

14 May 2026

A conversation with @sirupsen on scaling Shopify, building turbopuffer, and the future of databases. 0:00 - Scaling Shopify through flash sa…

ToolsDGX agent

A conversation with @sirupsen on scaling Shopify, building turbopuffer, and the future of databases. 0:00 - Scaling Shopify through flash sales and outages 8:13 - How top infrastructure teams collabor

Active Sensing with Meta-Reinforcement Learning for Emitter Localization from RF Observations

SafetyDGX agent

arXiv:2605.12569v1 Announce Type: cross Abstract: Global navigation satellite system (GNSS) interference poses a serious threat to reliable positioning, especially in indoor and multipath-rich environ

AdaptNC: Adaptive Nonconformity Scores for Conformal Prediction under Distribution Shift

SafetyDGX agent

arXiv:2602.01629v2 Announce Type: replace Abstract: Rigorous uncertainty quantification is essential for the safe deployment of autonomous systems in unconstrained environments. Conformal Prediction (

after 15 years of waiting, the developers of singapore gave up on waiting for the government to get the tech sector going and finally brough…

ToolsDGX agent

after 15 years of waiting, the developers of singapore gave up on waiting for the government to get the tech sector going and finally brought SF to SG. great showings from @daytonaio @usetusk @arizeai

Asynchronous Reasoning: Training-Free Interactive Thinking LLMs

SafetyDGX agent

arXiv:2512.10931v3 Announce Type: replace Abstract: Many state-of-the-art LLMs are trained to think before giving their answer. Reasoning can greatly improve language model capabilities, but it also m

Automated Rubrics for Reliable Evaluation of Medical Dialogue Systems

SafetyDGX agent

arXiv:2601.15161v2 Announce Type: replace-cross Abstract: Large Language Models (LLMs) are increasingly used for clinical decision support, where hallucinations and unsafe suggestions may pose direct

BEAVER: An Enterprise Benchmark for Text-to-SQL

Model ReleasesDGX agent

arXiv:2409.02038v3 Announce Type: replace-cross Abstract: Existing text-to-SQL benchmarks have largely been constructed from public databases with well-structured schemas and simplistic question-SQL p

BEHAVE: A Hybrid AI Framework for Real-Time Modeling of Collective Human Dynamics

SafetyDGX agent

arXiv:2605.12730v1 Announce Type: new Abstract: Existing AI systems for modeling human behavior operate at the level of individuals or detect events after they occur. As a result, they systematically

CodeClash: Benchmarking Goal-Oriented Software Engineering

Model ReleasesDGX agent

arXiv:2511.00839v2 Announce Type: replace-cross Abstract: Current benchmarks for coding evaluate language models (LMs) on concrete, well-specified tasks such as fixing specific bugs or writing targete

Contextual Bandits for Resource-Constrained Devices using Probabilistic Learning

Local AiDGX agent

arXiv:2605.13346v1 Announce Type: new Abstract: Contextual bandits (CB) are online sequential decision-making problems under partial feedback that underpin many adaptive services. There is a growing d

D-VLA: A High-Concurrency Distributed Asynchronous Reinforcement Learning Framework for Vision-Language-Action Models

Model ReleasesDGX agent

arXiv:2605.13276v1 Announce Type: new Abstract: The rapid evolution of Embodied AI has enabled Vision-Language-Action (VLA) models to excel in multimodal perception and task execution. However, applyi

Decoupling Exploration and Policy Optimization: Uncertainty Guided Tree Search for Hard Exploration

SafetyDGX agent

arXiv:2603.22273v4 Announce Type: replace Abstract: The process of discovery requires active exploration -- the act of collecting new and informative data. However, efficient autonomous exploration re

GRIP-VLM: Group-Relative Importance Pruning for Efficient Vision-Language Models

Local AiDGX agent

arXiv:2605.13375v1 Announce Type: cross Abstract: In Vision-Language Models (VLMs), processing a massive number of visual tokens incurs prohibitive computational overhead. While recent training-aware

HCSG: Human-Centric Semantic-Geometric Reasoning for Vision-Language Navigation

Model ReleasesDGX agent

arXiv:2605.13321v1 Announce Type: new Abstract: VLN has achieved remarkable progress by scaling data and model capacity. However, the assumption of a static environment breaks down in real-world indoo

Introducing Rime Mist v3 on Together AI, a production TTS family built for deterministic pronunciation and controllable voice output. AI nat…

ApplicationsDGX agent

Introducing Rime Mist v3 on Together AI, a production TTS family built for deterministic pronunciation and controllable voice output. AI natives can now deploy @rimelabs Mist v3 on Together AI dedicat

Large Language Models Lack Temporal Awareness of Medical Knowledge

Model ReleasesDGX agent

arXiv:2605.13045v1 Announce Type: new Abstract: The existing methods for evaluating the medical knowledge of Large Language Models (LLMs) are largely based on atemporal examination-style benchmarks, w

Limits of Personalizing Differential Privacy Budgets

ResearchDGX agent

arXiv:2605.13503v1 Announce Type: cross Abstract: A key technical difficulty in differential privacy is selecting a privacy budget that satisfies privacy requirements while maximizing utility. A natur

Prismatic World Model: Learning Compositional Dynamics for Planning in Hybrid Systems

ResearchDGX agent

arXiv:2512.08411v2 Announce Type: replace Abstract: Model-based planning in robotic domains is challenged by the hybrid nature of physical dynamics, where continuous motion is punctuated by discrete e

PROMETHEUS: Automating Deep Causal Research Integrating Text, Data and Models

Local AiDGX agent

arXiv:2605.12835v1 Announce Type: new Abstract: Large language models can extract local causal claims from text, but those claims become more useful when organized as persistent, navigable world model

Rigel3D: Rig-aware Latents for Animation-Ready 3D Asset Generation

ResearchDGX agent

arXiv:2605.13129v1 Announce Type: cross Abstract: Recent 3D generative models can synthesize high-quality assets, but their outputs are typically static: they lack the skeletal rigs, joint hierarchies

scShapeBench: Discovering geometry from high dimensional scRNAseq data

Model ReleasesDGX agent

arXiv:2605.12662v1 Announce Type: new Abstract: High-dimensional point cloud data arise across many scientific domains, especially single-cell biology. The shapes or topologies of these datasets deter

Senses Wide Shut: A Representation-Action Gap in Omnimodal LLMs

Model ReleasesDGX agent

arXiv:2605.13737v1 Announce Type: new Abstract: When an omnimodal large language model accepts a question whose textual premise contradicts what it actually sees or hears, does the failure lie in perc

SupChain-Bench: Benchmarking Large Language Models for Real-World Supply Chain Management

Model ReleasesDGX agent

arXiv:2602.07342v2 Announce Type: replace Abstract: Large language models (LLMs) have shown promise in complex reasoning and tool-based decision making, motivating their application to real-world supp

TiCo: Time-Controllable Spoken Dialogue Model

Model ReleasesDGX agent

arXiv:2603.22267v2 Announce Type: replace-cross Abstract: We introduce TiCo, a time-controllable spoken dialogue model (SDM) that follows time-constrained instructions (e.g., 'Please generate a respon

Together AI STT models now hold the top two spots for transcription speed on the @ArtificialAnlys Speech to Text leaderboard. NVIDIA Parakee…

HardwareDGX agent

Together AI STT models now hold the top two spots for transcription speed on the @ArtificialAnlys Speech to Text leaderboard. NVIDIA Parakeet TDT 0.6B V3 on Together AI ranks #1, transcribing 303 seco

Unweighted ranking for value-based decision making with uncertainty

SafetyDGX agent

arXiv:2605.13601v1 Announce Type: new Abstract: As intelligent systems are increasingly implemented in our society to make autonomous decisions, their commitment to human values raises serious concern

13 May 2026

A Survey of On-Policy Distillation for Large Language Models

SafetyDGX agent

arXiv:2604.00626v2 Announce Type: replace-cross Abstract: As Large Language Models (LLMs) continue to grow in both capability and cost, transferring frontier capabilities into smaller, deployable stud

我們開源了這顆星球🌎上速度最快的低成本 bm25 引擎。

Model ReleasesDGX agent

我們開源了這顆星球🌎上速度最快的低成本 bm25 引擎。 so we built psql_bm25s. exact BM25 retrieval. native Postgres access method. ~23x faster than pg_search on the standard benchmark. retrieval stops being a budget item. the

Characterizing the Robustness of Black-Box LLM Planners Under Perturbed Observations with Adaptive Stress Testing

SafetyDGX agent

arXiv:2505.05665v4 Announce Type: replace-cross Abstract: Large language models (LLMs) have recently demonstrated success in decision-making tasks including planning, control, and prediction, but thei

Entropy Polarity in Reinforcement Fine-Tuning: Direction, Asymmetry, and Control

SafetyDGX agent

arXiv:2605.11775v1 Announce Type: cross Abstract: Policy entropy has emerged as a fundamental measure for understanding and controlling exploration in reinforcement learning with verifiable rewards (R

Hierarchical LLM-Driven Control for HAPS-Assisted UAV Networks: Joint Optimization of Flight and Connectivity

Local AiDGX agent

arXiv:2605.11509v1 Announce Type: cross Abstract: Uncrewed aerial vehicles (UAVs) are increasingly deployed in complex networked environments, yet the joint optimization of multi-UAV motion control an

Intention-Conditioned Flow Occupancy Models

Model ReleasesDGX agent

arXiv:2506.08902v4 Announce Type: replace Abstract: Large-scale pre-training has fundamentally changed how machine learning research is done today: large foundation models are trained once, and then c

Internalizing Curriculum Judgment for LLM Reinforcement Fine-Tuning

SafetyDGX agent

arXiv:2605.11235v1 Announce Type: new Abstract: In LLM Reinforcement Fine-Tuning (RFT), curriculum learning drives both efficiency and performance. Yet, current methods externalize curriculum judgment

KV-Fold: One-Step KV-Cache Recurrence for Long-Context Inference

Model ReleasesDGX agent

arXiv:2605.12471v1 Announce Type: cross Abstract: We introduce KV-Fold, a simple, training-free long-context inference protocol that treats the key-value (KV) cache as the accumulator in a left fold o

Looking and Listening Inside and Outside: Multimodal Artificial Intelligence Systems for Driver Safety Assessment and Intelligent Vehicle Decision-Making

SafetyDGX agent

arXiv:2602.07668v2 Announce Type: replace Abstract: The looking-in-looking-out (LILO) framework has enabled intelligent vehicle applications that understand both the outside scene and the driver state

// δ-mem: Efficient Online Memory for LLMs // One of the more elegant memory mechanisms I've seen this month. Most long-term memory work eit…

TutorialsDGX agent

// δ-mem: Efficient Online Memory for LLMs // One of the more elegant memory mechanisms I've seen this month. Most long-term memory work either inflates context or retrains the model. This paper shows

PrivacySIM: Evaluating LLM Simulation of User Privacy Behavior

ApplicationsDGX agent

arXiv:2605.12147v1 Announce Type: cross Abstract: Large language models (LLMs) are increasingly used to simulate human behavior, but their ability to simulate individual privacy decisions is not well

Rainbow Deep Q-Learning with Kinematics-Aware Design for Cooperative Delta and 3-RRS Parallel Robot Insertion

SafetyDGX agent

arXiv:2605.11697v1 Announce Type: new Abstract: This paper presents a kinematics-aware deep reinforcement learning framework based on Rainbow Deep Q-Networks (DQN) for cooperative peg-in-hole manipula

SAGAS: Semantic-Aware Graph-Assisted Stitching for Offline Temporal Logic Planning

SafetyDGX agent

arXiv:2512.00775v2 Announce Type: replace Abstract: Linear Temporal Logic (LTL) provides a rigorous framework for specifying long-horizon robotic tasks, yet existing approaches face a trade-off: model

← Previous
1…278279280281282…294
Next →