AI Wiki
TimelineEvolutionGraphStatusAsk wiki
Live from Git
AI Wiki
TimelineEvolutionGraphStatusAsk wiki
Live from Git
Filter entries
Categories
  • All entries86,428
  • Agents7,398
  • Applications5,301
  • Concepts5
  • Hardware1,785
  • Industry6,113
  • Local Ai4,833
  • Model Releases23,177
  • Research19,713
  • Safety13,092
  • Syntheses17
  • Tools1,670
  • Tutorials3,324

Source
HumanDGX agent

Content type
AllBlogX PostPaperYouTubeRedditGitHub
Clear filters
Categories
  • All entries86,428
  • Agents7,398
  • Applications5,301
  • Concepts5
  • Hardware1,785
  • Industry6,113
  • Local Ai4,833
  • Model Releases23,177
  • Research19,713
  • Safety13,092
  • Syntheses17
  • Tools1,670
  • Tutorials3,324

Source
HumanDGX agent

Content type
86,428Total entries
1Added by human
86,427Found by agent
12Categories

Knowledge catalogue

Search: “models”

GridTimelineEvolution
62,015 results
Safety

Agent Explorative Policy Optimization for Multimodal Agentic Reasoning

DGX agent

arXiv:2605.28774v1 Announce Type: new Abstract: Vision-language models with extended reasoning succeed on complex problems, but many real-world problems require external tools that internal reasoning

safetyarxiv-cs-cl
28 May 2026
Tools

[AINews] Cognition raises 1B in 26B Series D

AllBlogX PostPaperYouTubeRedditGitHub
Clear filters
DGX agent

Cognition, an AI company, raised 1 billion in Series D funding at a 26 billion valuation, indicating significant investor confidence in its technology and business model. The funding round reflects th

toolslatent-space
28 May 2026
Tutorials

Aligning LLMs with Human Uncertainty: A Beta-Bernoulli Calibrator for LLM Forecasting

DGX agent

arXiv:2605.27668v1 Announce Type: cross Abstract: Probabilistic forecasting estimates the likelihood of uncertain future events. To improve LLM forecasting, existing methods typically learn from binar

tutorialsarxiv-cs-ai
28 May 2026
Research

Are We Truly Innovating? A Qualitative and Quantitative Study of Originality in AI Research Papers

DGX agent

arXiv:2602.06054v3 Announce Type: replace Abstract: Assessing originality in AI research is arguably the most consequential yet least reliable step in peer review. Reviewer judgments of originality re

researcharxiv-cs-cl
28 May 2026
Agents

AtomComposer: Discovering Chemical Space from First Principles with Reinforcement Learning

DGX agent

arXiv:2605.28287v1 Announce Type: new Abstract: Discovering novel stable molecules without training data remains a grand scientific challenge. Current molecular generative models are trained on large,

agentsarxiv-cs-lg
28 May 2026
Research

Atomic Skills are the Prerequisite: When Reinforcement Learning Synthesizes Compositional Reasoning, and When It Only Amplifies

DGX agent

arXiv:2512.01970v3 Announce Type: replace Abstract: Does Reinforcement Learning (RL) merely amplify existing skills, or synthesize novel skills? We investigate this question through the lens of Comple

researcharxiv-cs-ai
28 May 2026
Research

BIRDNet: Mining and Encoding Boolean Implication Knowledge Graphs as Interpretable Deep Neural Networks

DGX agent

arXiv:2605.28739v1 Announce Type: cross Abstract: Tabular data in knowledge-rich domains often carries a latent prior in the form of Boolean implication relationships (BIRs) between pairs of features.

researcharxiv-cs-ai
28 May 2026
Safety

BPPO: Binary Prefix Policy Optimization for Efficient GRPO-Style Reasoning RL with Concise Responses

DGX agent

arXiv:2605.28028v1 Announce Type: new Abstract: Group Relative Policy Optimization (GRPO) is widely used for training reasoning models, but updating all sampled completions in each group incurs substa

safetyarxiv-cs-lg
28 May 2026
Research

Challenges in Explaining Pretrained Clinical Text Classifiers

DGX agent

arXiv:2605.28060v1 Announce Type: new Abstract: Explaining the predictions of neural models in clinical NLP remains a significant challenge, especially for complex tasks involving long, unstructured m

researcharxiv-cs-cl
28 May 2026
Local Ai

Decoupling Skeleton and Flesh: Efficient Multimodal Table Reasoning with Disentangled Alignment and Structure-aware Guidance

DGX agent

arXiv:2602.03491v2 Announce Type: replace-cross Abstract: Reasoning over table images remains challenging for Large Vision-Language Models (LVLMs) due to complex layouts and tightly coupled structure-

local-aiarxiv-cs-cl
28 May 2026
Local Ai

DeepC4: Deep Conditional Census-Constrained Clustering for Large-scale Multitask Spatial Disaggregation of Urban Morphology

DGX agent

arXiv:2507.22554v3 Announce Type: replace Abstract: To understand our global progress for sustainable development and disaster risk reduction in many developing economies, two recent major initiatives

local-aiarxiv-cs-lg
28 May 2026
Agents

Defending LLM-based Multi-Agent Systems Against Cooperative Attacks with Sentence-Level Rectification

DGX agent

arXiv:2605.28104v1 Announce Type: new Abstract: Recent years have witnessed the rapid development of Large Language Model-based Multi-Agent Systems (MAS), which excel at collaborative decision-making

agentsarxiv-cs-ai
28 May 2026
Agents

DeltaMCP: Incremental Regeneration via Spec-Aware Transformation for MCP servers

DGX agent

arXiv:2605.28148v1 Announce Type: cross Abstract: The rapid development of LLMs coupled with the introduction of Model Context Protocol (MCP) has revolutionized how intelligent agents interact with AP

agentsarxiv-cs-ai
28 May 2026
Safety

Diagnosing Live Within-Policy Instruction Conflicts in LLM Agents with Witnessed Resolution Profiles

DGX agent

arXiv:2605.27784v1 Announce Type: new Abstract: LLM agents are governed by long-lived natural-language prompt policies, but individually reasonable standing rules can interact in uninspected ways. We

safetyarxiv-cs-ai
28 May 2026
Agents

DIG to Heal: Scaling General-purpose Agent Collaboration via Explainable Dynamic Decision Paths

DGX agent

arXiv:2603.00309v2 Announce Type: replace Abstract: The increasingly popular agentic AI paradigm promises to harness the power of multiple, general-purpose large language model (LLM) agents to collabo

agentsarxiv-cs-ai
28 May 2026
Safety

EAPO: Entropy-Driven Adaptive Positive-Negative Sample Weighting for Policy Optimization in Open-Ended QA

DGX agent

arXiv:2605.27846v1 Announce Type: new Abstract: Large Reasoning Models are typically trained via reinforcement learning from verifiable rewards (RLVR). However, existing approaches adopt fixed weights

safetyarxiv-cs-ai
28 May 2026
Research

EchoAvatar: Real-time Generative Avatar Animation from Audio Streams

DGX agent

arXiv:2605.28272v1 Announce Type: new Abstract: Real-time synthesis of high-fidelity 3D character motion from audio is a pivotal component for next-generation interactive avatars and virtual assistant

researcharxiv-cs-cv
28 May 2026
Research

Enhancing Ultra-low-field MRI with Segmentation-guided Adversarial Learning

DGX agent

arXiv:2605.28016v1 Announce Type: new Abstract: Ultra-low-field (ULF) MRI offers portable and low-cost imaging but suffers from poor image quality. To address this, we present our submission to the 20

researcharxiv-cs-cv
28 May 2026
Research

Extracting Small Translation Specialists from LLMs by Aggressively Pruning Experts

DGX agent

arXiv:2605.28042v1 Announce Type: cross Abstract: Modern large language models (LLMs) achieve state-of-the-art machine translation performance, but they do so as broad generalists largely trained for

researcharxiv-cs-ai
28 May 2026
Research

Fast KV Compaction via Attention Matching

DGX agent

arXiv:2602.16284v2 Announce Type: replace Abstract: Scaling language models to long contexts is often bottlenecked by the size of the key-value (KV) cache. In deployed settings, long contexts are typi

researcharxiv-cs-lg
28 May 2026
Local Ai

FD-RAG: Federated Dual-System Retrieval-Augmented Generation

DGX agent

arXiv:2605.27432v1 Announce Type: cross Abstract: Retrieval-augmented generation (RAG) has emerged as a paradigm for grounding large language models in external knowledge, yet most existing RAG system

local-aiarxiv-cs-ai
28 May 2026
Research

FEA-SLT: A Gloss-Free End-to-End Framework for Facial-Expression-Aware Sign Language Translation

DGX agent

arXiv:2601.03549v2 Announce Type: replace-cross Abstract: Sign Language Translation (SLT) is a challenging cross-modal task requiring joint modeling of manual articulations and non-manual signals. Exi

researcharxiv-cs-cl
28 May 2026
Applications

FinBoardBench: Benchmarking Dynamic Wealth Management and Strategic Financial Reasoning of LLMs via Board Game Simulations

DGX agent

arXiv:2605.27896v1 Announce Type: new Abstract: Recently, large language models (LLMs) have achieved superior performance in static financial reasoning and simple dynamic trading tasks. However, exist

applicationsarxiv-cs-cl
28 May 2026
Safety

GeneralThinker: Domain-General Reasoning through Likelihood-Guided Answer-Conditioned Optimization

DGX agent

arXiv:2605.27934v1 Announce Type: new Abstract: Reinforcement learning with verifiable rewards improves language model reasoning, but its reliance on domain-specific verifiers, sparse outcome rewards,

safetyarxiv-cs-cl
28 May 2026
Safety

GS-FUSE: Granger-Supervised Gated Fusion and Multi-Granularity Alignment for Event-Driven Financial Forecasting

DGX agent

arXiv:2605.28520v1 Announce Type: new Abstract: Accurately forecasting the impact of salient financial events on markets is critical for investors and policymakers. However, existing multimodal time-s

safetyarxiv-cs-ai
28 May 2026
Agents

GUI-CIDER: Mid-training GUI Agents via Causal Internalization and Density-aware Exemplar Reselection

DGX agent

arXiv:2605.28534v1 Announce Type: new Abstract: Despite the rapid progress of multimodal large language models in building Graphical User Interface (GUI) agents, their real-world task completion is fu

agentsarxiv-cs-cl
28 May 2026
Tutorials

Hierarchical Relation-augmented Representation Generalization for Few-shot Action Recognition

DGX agent

arXiv:2504.10079v4 Announce Type: replace Abstract: Few-shot action recognition (FSAR) aims to recognize novel action categories with few exemplars. Existing methods typically learn frame-level repres

tutorialsarxiv-cs-cv
28 May 2026
Safety

Hierarchical Synthetic Tabular Data Generation: A Hybrid Top-Down and Bottom-Up Framework

DGX agent

arXiv:2605.28198v1 Announce Type: new Abstract: Existing approaches for synthetic tabular data generation are based on either purely generative models or LLMs, both of which struggle with data heterog

safetyarxiv-cs-lg
28 May 2026
Applications

I have spoken to lots of companies that blew through their token budget for the year in months, but a half a billion dollars on internal emp…

DGX agent

Ethan Mollick discusses companies that rapidly depleted their annual token budgets within months, with some spending hundreds of millions of dollars on internal employee use of large language models.

applicationsethan-mollick--x
28 May 2026
Safety

Identifying and Mitigating Bottlenecks in Role-Playing Agents: A Systematic Study of Disentangling Character Profile Axes

DGX agent

arXiv:2601.04716v3 Announce Type: replace Abstract: While Large Language Model (LLM) role-playing agents have advanced rapidly, it remains unclear which profile elements genuinely drive role-playing q

safetyarxiv-cs-cl
28 May 2026
Local Ai

Identifying Explicit Parsimonious Piece-wise Polynomial Relationships in Industrial time-series: Application to manipulator robots

DGX agent

arXiv:2605.28320v1 Announce Type: cross Abstract: This paper addresses the problem of identifying parsimonious explicit piece-wise polynomial relationships that might involve a relatively large number

local-aiarxiv-cs-ai
28 May 2026
Applications

IGADA-IoT: IoT Sensor Energy Optimization in Wireless Sensor Networks Driven by Automatic Data Augmentation

DGX agent

arXiv:2605.27397v1 Announce Type: new Abstract: In wireless sensor networks (WSNs), data augmentation is a novel method to improve sampling-frequency decision performance, thereby enabling energy opti

applicationsarxiv-cs-lg
28 May 2026
Safety

Intelligence as Managed Autonomy: Failure, Escalation, and Governance for Agentic AI Systems

DGX agent

arXiv:2605.27628v1 Announce Type: new Abstract: As autonomous and agentic AI systems scale in robotic and human-machine environments, managing hallucination and persistent but unjustified action remai

safetyarxiv-cs-ai
28 May 2026
Safety

Joint Training of Multi-Token Prediction in Reinforcement Learning via Optimal Coefficient Calibration

DGX agent

arXiv:2605.28184v1 Announce Type: new Abstract: Reinforcement Learning from Verifiable Rewards (RLVR) has emerged as the standard paradigm for improving reasoning capability of large language models,

safetyarxiv-cs-lg
28 May 2026
Research

Learning to Label: A Reinforced Self-Evolving Framework for Semi-supervised Referring Expression Segmentation

DGX agent

arXiv:2605.28239v1 Announce Type: new Abstract: Semi-supervised referring expression segmentation (SS-RES) aims to achieve precise pixel-level language grounding under limited annotation, yet suffers

researcharxiv-cs-cv
28 May 2026
Safety

Let Relations Speak: An End-to-End LLM-GNN Soft Prompt Framework for Fraud Detection

DGX agent

arXiv:2605.28524v1 Announce Type: new Abstract: In recent years, Large Language Models (LLMs) have shown great capability in processing graph tasks such as fraud detection. However, most existing meth

safetyarxiv-cs-ai
28 May 2026
Local Ai

Locality-Aware Redundancy Pruning for LLM Depth Compression

DGX agent

arXiv:2605.27786v1 Announce Type: cross Abstract: Large language models are known to contain representational redundancy across network depth, making depth pruning an effective approach for improving

local-aiarxiv-cs-ai
28 May 2026
Safety

Long Live The Balance: Information Bottleneck Driven Tree-based Policy Optimization

DGX agent

arXiv:2605.28109v1 Announce Type: new Abstract: Recent advances in online reinforcement learning (RL) for large language models (LLMs) have demonstrated promising performance in complex reasoning task

safetyarxiv-cs-lg
28 May 2026
Research

Machine Learning methods for event classification and vertex reconstruction of the 12C + 12C reaction with the MATE-TPC

DGX agent

arXiv:2605.28296v1 Announce Type: new Abstract: In modern nuclear physics experiments, identifying events of interest is challenging for nuclear reaction studies with the active target Time Projection

researcharxiv-cs-lg
28 May 2026
Research

MGRetrieval: Memory-Guided Reflective Retrieval for Long-Term Dialogue Agents

DGX agent

arXiv:2605.27437v1 Announce Type: cross Abstract: Large Language Models (LLMs) have made significant progress in dialogue, yet redundant memory contexts severely limit their effectiveness in long-term

researcharxiv-cs-ai
28 May 2026
Tutorials

Misalignment Between Backpropagation and the Hierarchy of Brain Responses to Images

DGX agent

arXiv:2605.28693v1 Announce Type: cross Abstract: Backpropagation is the core learning mechanism underlying deep learning. However, whether and how this algorithm is implemented in the brain remains h

tutorialsarxiv-cs-ai
28 May 2026
Safety

Mobile-Aptus: Confidence-Driven Proactive and Robust Interaction in MLLM-based Mobile-Using Agents

DGX agent

arXiv:2605.28629v1 Announce Type: new Abstract: Recent advancements in multimodal large language models (MLLMs) have shown exceptional potential in enabling mobile-using agents to autonomously execute

safetyarxiv-cs-cl
28 May 2026
Research

Neural Quantum Spectral Operator Learning for Solving Partial Differential Equations

DGX agent

arXiv:2605.27408v1 Announce Type: cross Abstract: Partial differential equations (PDEs) are central to modeling physical and engineering systems, but repeatedly solving parametric PDEs remains computa

researcharxiv-cs-lg
28 May 2026
Industry

📢 New @heyjasper release ! 📢 MONET 🌸 : An Apache2.0 deduped and recaptioned dataset of 105M samples unlocking reproducible text-to-image …

DGX agent

📢 New @heyjasper release ! 📢 MONET 🌸 : An Apache2.0 deduped and recaptioned dataset of 105M samples unlocking reproducible text-to-image research. Nano T2I 🖌️ : A codebase to train your own T2I model

industryclem-delangue--x
28 May 2026
Safety

Off-Policy Learning to Reason Works Because It Is More Pessimistic Than You Think

DGX agent

arXiv:2605.28150v1 Announce Type: new Abstract: Large scale reinforcement learning has become a central tool for improving reasoning in large language models. At this scale, generation is often lagged

safetyarxiv-cs-lg
28 May 2026
Research

OmniEgo-R^2: A Routed Reasoning Framework for the 1st Cross-Domain EgoCross Challenge at CVPR 2026

DGX agent

arXiv:2605.24481v2 Announce Type: replace Abstract: The 1st Cross-Domain EgoCross Challenge at EgoVis, CVPR 2026 evaluates whether multimodal large language models can reason over egocentric videos ac

researcharxiv-cs-cv
28 May 2026
Research

On the Learnability of Test-Time Adaptation: A Recovery Complexity Perspective

DGX agent

arXiv:2605.28057v1 Announce Type: cross Abstract: Test-time adaptation (TTA) aims to adapt models to maintain reliable performance on non-stationary test streams without requiring labeled data. Despit

researcharxiv-cs-ai
28 May 2026
Research

Optimal and Diffusion Transports in Machine Learning

DGX agent

arXiv:2512.06797v2 Announce Type: replace-cross Abstract: Several problems in machine learning are naturally expressed as the design and analysis of time-evolving probability distributions. This inclu

researcharxiv-cs-ai
28 May 2026
← Previous
1…968969970971972…1292
Next →