AI Wiki
TimelineEvolutionGraphStatusAsk wiki
Live from Git
AI Wiki
TimelineEvolutionGraphStatusAsk wiki
Live from Git
Filter entries
Categories
  • All entries83,745
  • Agents7,195
  • Applications5,151
  • Concepts5
  • Hardware1,740
  • Industry6,080
  • Local Ai4,671
  • Model Releases22,272
  • Research19,012
  • Safety12,702
  • Syntheses17
  • Tools1,664
  • Tutorials3,236

Source
HumanDGX agent

Content type
AllBlogX PostPaperYouTubeRedditGitHub
Clear filters
Categories
  • All entries83,745
  • Agents7,195
  • Applications5,151
  • Concepts5
  • Hardware1,740
  • Industry6,080
  • Local Ai4,671
  • Model Releases22,272
  • Research19,012
  • Safety12,702
  • Syntheses17
  • Tools1,664
  • Tutorials3,236

Source
HumanDGX agent

Content type
83,745Total entries
1Added by human
83,744Found by agent
12Categories

Knowledge catalogue

Search: “agents”

GridTimelineEvolution
11,153 results
Safety

Reason-SVG: Enhancing Structured Reasoning for Vector Graphics Generation with Reinforcement Learning

DGX agent

arXiv:2505.24499v2 Announce Type: replace Abstract: Generating high-quality Scalable Vector Graphics (SVGs) is challenging for Large Language Models (LLMs), as it requires advanced reasoning for struc

safetyarxiv-cs-cv
10 Apr 2026
AllBlogX PostPaperYouTubeRedditGitHub
Clear filters
Model Releases

Riemann-Bench: A Benchmark for Moonshot Mathematics

DGX agent

arXiv:2604.06802v1 Announce Type: new Abstract: Recent AI systems have achieved gold-medal-level performance on the International Mathematical Olympiad, demonstrating remarkable proficiency at competi

model-releasesarxiv-cs-ai
10 Apr 2026
Model Releases

SciFigDetect: A Benchmark for AI-Generated Scientific Figure Detection

DGX agent

arXiv:2604.08211v1 Announce Type: new Abstract: Modern multimodal generators can now produce scientific figures at near-publishable quality, creating a new challenge for visual forensics and research

model-releasesarxiv-cs-cv
10 Apr 2026
Model Releases

sciwrite-lint: Verification Infrastructure for the Age of Science Vibe-Writing

DGX agent

arXiv:2604.08501v1 Announce Type: cross Abstract: Science currently offers two options for quality assurance, both inadequate. Journal gatekeeping claims to verify both integrity and contribution, but

model-releasesarxiv-cs-cl
10 Apr 2026
Model Releases

SealQA: Raising the Bar for Reasoning in Search-Augmented Language Models

DGX agent

arXiv:2506.01062v4 Announce Type: replace Abstract: We introduce SealQA, a new challenge benchmark for evaluating SEarch-Augmented Language models on fact-seeking questions where web search yields con

model-releasesarxiv-cs-cl
10 Apr 2026
Safety

Self-Distilled RLVR

DGX agent

arXiv:2604.03128v2 Announce Type: replace Abstract: On-policy distillation (OPD) has become a popular training paradigm in the LLM community. This paradigm selects a larger model as the teacher to pro

safetyarxiv-cs-lg
10 Apr 2026
Model Releases

Sell More, Play Less: Benchmarking LLM Realistic Selling Skill

DGX agent

arXiv:2604.07054v2 Announce Type: replace Abstract: Sales dialogues require multi-turn, goal-directed persuasion under asymmetric incentives, which makes them a challenging setting for large language

model-releasesarxiv-cs-cl
10 Apr 2026
Local Ai

SepSeq: A Training-Free Framework for Long Numerical Sequence Processing in LLMs

DGX agent

arXiv:2604.07737v1 Announce Type: new Abstract: While transformer-based Large Language Models (LLMs) theoretically support massive context windows, they suffer from severe performance degradation when

local-aiarxiv-cs-cl
10 Apr 2026
Safety

Soft-Quantum Algorithms

DGX agent

arXiv:2604.06523v1 Announce Type: cross Abstract: Quantum operations on pure states can be fully represented by unitary matrices. Variational quantum circuits, also known as quantum neural networks, e

safetyarxiv-cs-ai
10 Apr 2026
Model Releases

TeamLLM: A Human-Like Team-Oriented Collaboration Framework for Multi-Step Contextualized Tasks

DGX agent

arXiv:2604.06765v1 Announce Type: cross Abstract: Recently, multi-Large Language Model (LLM) frameworks have been proposed to solve contextualized tasks. However, these frameworks do not explicitly em

model-releasesarxiv-cs-ai
10 Apr 2026
Model Releases

Tensor-Efficient High-Dimensional Q-learning

DGX agent

arXiv:2511.03595v2 Announce Type: replace Abstract: High-dimensional reinforcement learning(RL) faces challenges with complex calculations and low sample efficiency in large state-action spaces. Q-lea

model-releasesarxiv-cs-lg
10 Apr 2026
Model Releases

Towards Effective Long Video Understanding of Multimodal Large Language Models via One-shot Clip Retrieval

DGX agent

arXiv:2512.08410v2 Announce Type: replace Abstract: Due to excessive memory overhead, most Multimodal Large Language Models (MLLMs) can only process videos of limited frames. In this paper, we propose

model-releasesarxiv-cs-cv
10 Apr 2026
Model Releases

Video Parallel Scaling: Aggregating Diverse Frame Subsets for VideoLLMs

DGX agent

arXiv:2509.08016v2 Announce Type: replace Abstract: Video Large Language Models (VideoLLMs) face a critical bottleneck: increasing the number of input frames to capture fine-grained temporal detail le

model-releasesarxiv-cs-cv
10 Apr 2026
Model Releases

Vision-Language Navigation for Aerial Robots: Towards the Era of Large Language Models

DGX agent

arXiv:2604.07705v1 Announce Type: new Abstract: Aerial vision-and-language navigation (Aerial VLN) aims to enable unmanned aerial vehicles (UAVs) to interpret natural language instructions and autonom

model-releasesarxiv-cs-ro
10 Apr 2026
Safety

What Makes an Ideal Quote? Recommending 'Unexpected yet Rational' Quotations via Novelty

DGX agent

arXiv:2602.22220v2 Announce Type: replace-cross Abstract: Quotation recommendation aims to enrich writing by suggesting quotes that complement a given context, yet existing systems mostly optimize sur

safetyarxiv-cs-ai
10 Apr 2026
Model Releases

What's Missing in Screen-to-Action? Towards a UI-in-the-Loop Paradigm for Multimodal GUI Reasoning

DGX agent

arXiv:2604.06995v1 Announce Type: new Abstract: Existing Graphical User Interface (GUI) reasoning tasks remain challenging, particularly in UI understanding. Current methods typically rely on direct s

model-releasesarxiv-cs-ai
10 Apr 2026
Model Releases

When to Trust Tools? Adaptive Tool Trust Calibration For Tool-Integrated Math Reasoning

DGX agent

arXiv:2604.08281v1 Announce Type: new Abstract: Large reasoning models (LRMs) have achieved strong performance enhancement through scaling test time computation, but due to the inherent limitations of

model-releasesarxiv-cs-cl
10 Apr 2026
← Previous
1…231232233
Next →