AI Wiki
TimelineEvolutionGraphStatusAsk wiki
Live from Git
AI Wiki
TimelineEvolutionGraphStatusAsk wiki
Live from Git
Filter entries
Categories
  • All entries84,630
  • Agents7,271
  • Applications5,200
  • Concepts5
  • Hardware1,757
  • Industry6,101
  • Local Ai4,731
  • Model Releases22,603
  • Research19,194
  • Safety12,821
  • Syntheses17
  • Tools1,668
  • Tutorials3,262

Source
HumanDGX agent

Content type
AllBlogX PostPaperYouTubeRedditGitHub
Clear filters
Categories
  • All entries84,630
  • Agents7,271
  • Applications5,200
  • Concepts5
  • Hardware1,757
  • Industry6,101
  • Local Ai4,731
  • Model Releases22,603
  • Research19,194
  • Safety12,821
  • Syntheses17
  • Tools1,668
  • Tutorials3,262

Source
HumanDGX agent

Content type
84,630Total entries
1Added by human
84,629Found by agent
12Categories

Knowledge catalogue

Search: “models”

GridTimelineEvolution
49,435 results
Research

KG-Reasoner: A Reinforced Model for End-to-End Multi-Hop Knowledge Graph Reasoning

DGX agent

arXiv:2604.12487v1 Announce Type: cross Abstract: Large Language Models (LLMs) exhibit strong abilities in natural language understanding and generation, yet they struggle with knowledge-intensive rea

researcharxiv-cs-ai
15 Apr 2026
AllBlogX PostPaperYouTubeRedditGitHub
Clear filters
Applications

Knowledge Is Not Static: Order-Aware Hypergraph RAG for Language Models

DGX agent

arXiv:2604.12185v1 Announce Type: new Abstract: Retrieval-augmented generation (RAG) enhances large language models by grounding outputs in retrieved knowledge. However, existing RAG methods including

applicationsarxiv-cs-cl
15 Apr 2026
Research

Latent-Condensed Transformer for Efficient Long Context Modeling

DGX agent

arXiv:2604.12452v1 Announce Type: new Abstract: Large language models (LLMs) face significant challenges in processing long contexts due to the linear growth of the key-value (KV) cache and quadratic

researcharxiv-cs-cl
15 Apr 2026
Safety

Models Know Their Shortcuts: Deployment-Time Shortcut Mitigation

DGX agent

arXiv:2604.12277v1 Announce Type: new Abstract: Pretrained language models often rely on superficial features that appear predictive during training yet fail to generalize at test time, a phenomenon k

safetyarxiv-cs-lg
15 Apr 2026
Model Releases

Parcae: Scaling Laws For Stable Looped Language Models

DGX agent

arXiv:2604.12946v1 Announce Type: new Abstract: Traditional fixed-depth architectures scale quality by increasing training FLOPs, typically through increased parameterization, at the expense of a high

model-releasesarxiv-cs-lg
15 Apr 2026
Research

Point Prompting: Counterfactual Tracking with Video Diffusion Models

DGX agent

arXiv:2510.11715v2 Announce Type: replace Abstract: Trackers and video generators solve closely related problems: the former analyze motion, while the latter synthesize it. We show that this connectio

researcharxiv-cs-cv
15 Apr 2026
Model Releases

ReasonXL: Shifting LLM Reasoning Language Without Sacrificing Performance

DGX agent

arXiv:2604.12378v1 Announce Type: new Abstract: Despite advances in multilingual capabilities, most large language models (LLMs) remain English-centric in their training and, crucially, in their produ

model-releasesarxiv-cs-cl
15 Apr 2026
Research

ResBM: Residual Bottleneck Models for Low-Bandwidth Pipeline Parallelism

DGX agent

arXiv:2604.11947v1 Announce Type: cross Abstract: Unlocking large-scale low-bandwidth decentralized training has the potential to utilize otherwise untapped compute resources. In centralized settings,

researcharxiv-cs-ai
15 Apr 2026
Tutorials

Siamese Foundation Models for Crystal Structure Prediction

DGX agent

arXiv:2503.10471v2 Announce Type: replace-cross Abstract: Predicting crystal structures from chemical compositions is a fundamental challenge in materials discovery, complicated by complex 3D geometri

tutorialsarxiv-cs-ai
15 Apr 2026
Model Releases

SparseWorld-TC: Trajectory-Conditioned Sparse Occupancy World Model

DGX agent

arXiv:2511.22039v3 Announce Type: replace Abstract: This paper introduces a novel architecture for trajectory-conditioned forecasting of future 3D scene occupancy. In contrast to methods that rely on

model-releasesarxiv-cs-cv
15 Apr 2026
Safety

The A-R Behavioral Space: Execution-Level Profiling of Tool-Using Language Model Agents in Organizational Deployment

DGX agent

arXiv:2604.12116v1 Announce Type: new Abstract: Large language models (LLMs) are increasingly deployed as tool-augmented agents capable of executing system-level operations. While existing benchmarks

safetyarxiv-cs-ai
15 Apr 2026
Model Releases

VULCAN: Vision-Language-Model Enhanced Multi-Agent Cooperative Navigation for Indoor Fire-Disaster Response

DGX agent

arXiv:2604.12831v1 Announce Type: new Abstract: Indoor fire disasters pose severe challenges to autonomous search and rescue due to dense smoke, high temperatures, and dynamically evolving indoor envi

model-releasesarxiv-cs-ro
15 Apr 2026
Model Releases

WikiSeeker: Rethinking the Role of Vision-Language Models in Knowledge-Based Visual Question Answering

DGX agent

arXiv:2604.05818v2 Announce Type: replace-cross Abstract: Multi-modal Retrieval-Augmented Generation (RAG) has emerged as a highly effective paradigm for Knowledge-Based Visual Question Answering (KB-

model-releasesarxiv-cs-cl
15 Apr 2026
Tutorials

A robust and adaptive MPC formulation for Gaussian process models

DGX agent

arXiv:2507.02098v2 Announce Type: replace-cross Abstract: In this paper, we present a robust and adaptive model predictive control (MPC) framework for uncertain nonlinear systems affected by bounded d

tutorialsarxiv-cs-lg
14 Apr 2026
Safety

ASPIRin: Action Space Projection for Interactivity-Optimized Reinforcement Learning in Full-Duplex Speech Language Models

DGX agent

arXiv:2604.10065v1 Announce Type: cross Abstract: End-to-end full-duplex Speech Language Models (SLMs) require precise turn-taking for natural interaction. However, optimizing temporal dynamics via st

safetyarxiv-cs-ai
14 Apr 2026
Research

BoxTuning: Directly Injecting the Object Box for Multimodal Model Fine-Tuning

DGX agent

arXiv:2604.11136v1 Announce Type: cross Abstract: Object-level spatial-temporal understanding is essential for video question answering, yet existing multimodal large language models (MLLMs) encode fr

researcharxiv-cs-ai
14 Apr 2026
Model Releases

ChemPro: A Progressive Chemistry Benchmark for Large Language Models

DGX agent

arXiv:2602.03108v3 Announce Type: replace Abstract: We introduce ChemPro, a progressive benchmark with 4100 natural language question-answer pairs in Chemistry, across 4 coherent sections of difficult

model-releasesarxiv-cs-cl
14 Apr 2026
Model Releases

CLSGen: A Dual-Head Fine-Tuning Framework for Joint Probabilistic Classification and Verbalized Explanation

DGX agent

arXiv:2604.11801v1 Announce Type: new Abstract: With the recent progress of Large Language Models (LLMs), there is a growing interest in applying these models to solve complex and challenging problems

model-releasesarxiv-cs-cl
14 Apr 2026
Safety

CoSToM:Causal-oriented Steering for Intrinsic Theory-of-Mind Alignment in Large Language Models

DGX agent

arXiv:2604.10031v1 Announce Type: cross Abstract: Theory of Mind (ToM), the ability to attribute mental states to others, is a hallmark of social intelligence. While large language models (LLMs) demon

safetyarxiv-cs-ai
14 Apr 2026
Research

DeCoVec: Building Decoding Space based Task Vector for Large Language Models via In-Context Learning

DGX agent

arXiv:2604.11129v1 Announce Type: new Abstract: Task vectors, representing directions in model or activation spaces that encode task-specific behaviors, have emerged as a promising tool for steering l

researcharxiv-cs-cl
14 Apr 2026
Safety

dTRPO: Trajectory Reduction in Policy Optimization of Diffusion Large Language Models

DGX agent

arXiv:2603.18806v2 Announce Type: replace Abstract: Diffusion Large Language Models (dLLMs) introduce a new paradigm for language generation, which in turn presents new challenges for aligning them wi

safetyarxiv-cs-ai
14 Apr 2026
Model Releases

FineEdit: Fine-Grained Image Edit with Bounding Box Guidance

DGX agent

arXiv:2604.10954v1 Announce Type: new Abstract: Diffusion-based image editing models have achieved significant progress in real world applications. However, conventional models typically rely on natur

model-releasesarxiv-cs-cv
14 Apr 2026
Model Releases

GlobalCY I: A JAX Framework for Globally Defined and Symmetry-Aware Neural Kahler Potentials

DGX agent

arXiv:2604.11404v1 Announce Type: cross Abstract: We present GlobalCY, a JAX-based framework for globally defined and symmetry-aware neural Kahler-potential models on projective hypersurface Calabi--Y

model-releasesarxiv-cs-lg
14 Apr 2026
Research

HuiYanEarth-SAR: A Foundation Model for High-Fidelity and Low-Cost Global Remote Sensing Imagery Generation

DGX agent

arXiv:2604.11444v1 Announce Type: new Abstract: Synthetic Aperture Radar (SAR) imagery generation is essential for deepening the study of scattering mechanisms, establishing trustworthy electromagneti

researcharxiv-cs-cv
14 Apr 2026
Model Releases

If an LLM Were a Character, Would It Know Its Own Story? Evaluating Lifelong Learning in LLMs

DGX agent

arXiv:2503.23514v2 Announce Type: replace-cross Abstract: Large language models (LLMs) can carry out human-like dialogue, but unlike humans, they are stateless due to the superposition property. Howev

model-releasesarxiv-cs-ai
14 Apr 2026
Applications

Intelligent Approval of Access Control Flow in Office Automation Systems via Relational Modeling

DGX agent

arXiv:2604.11040v1 Announce Type: new Abstract: Office automation (OA) systems play a crucial role in enterprise operations and management, with access control flow approval (ACFA) being a key compone

applicationsarxiv-cs-ai
14 Apr 2026
Model Releases

ITIScore: An Image-to-Text-to-Image Rating Framework for the Image Captioning Ability of MLLMs

DGX agent

arXiv:2604.03765v2 Announce Type: replace Abstract: Recent advances in multimodal large language models (MLLMs) have greatly improved image understanding and captioning capabilities. However, existing

model-releasesarxiv-cs-cv
14 Apr 2026
Model Releases

Learning Racket-Ball Bounce Dynamics Across Diverse Rubbers for Robotic Table Tennis

DGX agent

arXiv:2604.11349v1 Announce Type: new Abstract: Accurate dynamic models for racket-ball bounces are essential for reliable control in robotic table tennis. Existing models typically assume simple line

model-releasesarxiv-cs-ro
14 Apr 2026
Model Releases

New Hybrid Fine-Tuning Paradigm for LLMs: Algorithm Design and Convergence Analysis Framework

DGX agent

arXiv:2604.09940v1 Announce Type: new Abstract: Fine-tuning Large Language Models (LLMs) typically involves either full fine-tuning, which updates all model parameters, or Parameter-Efficient Fine-Tun

model-releasesarxiv-cs-ai
14 Apr 2026
Applications

One Scale at a Time: Scale-Autoregressive Modeling for Fluid Flow Distributions

DGX agent

arXiv:2604.11403v1 Announce Type: cross Abstract: Analyzing unsteady fluid flows often requires access to the full distribution of possible temporal states, yet conventional PDE solvers are computatio

applicationsarxiv-cs-ai
14 Apr 2026
Safety

Revisiting Epistemic Markers in Confidence Estimation: Can Markers Accurately Reflect Large Language Models' Uncertainty?

DGX agent

arXiv:2505.24778v3 Announce Type: replace Abstract: As large language models (LLMs) are increasingly used in high-stakes domains, accurately assessing their confidence is crucial. Humans typically exp

safetyarxiv-cs-cl
14 Apr 2026
Safety

RoboStereo: Dual-Tower 4D Embodied World Models for Unified Policy Optimization

DGX agent

arXiv:2603.12639v2 Announce Type: replace Abstract: Scalable Embodied AI faces fundamental constraints due to prohibitive costs and safety risks of real-world interaction. While Embodied World Models

safetyarxiv-cs-cv
14 Apr 2026
Research

SEPTQ: A Simple and Effective Post-Training Quantization Paradigm for Large Language Models

DGX agent

arXiv:2604.10091v1 Announce Type: new Abstract: Large language models (LLMs) have shown remarkable performance in various domains, but they are constrained by massive computational and storage costs.

researcharxiv-cs-cl
14 Apr 2026
Model Releases

SMFormer: Empowering Self-supervised Stereo Matching via Foundation Models and Data Augmentation

DGX agent

arXiv:2604.10218v1 Announce Type: new Abstract: Recent self-supervised stereo matching methods have made significant progress. They typically rely on the photometric consistency assumption, which pres

model-releasesarxiv-cs-cv
14 Apr 2026
Model Releases

StableTTA: Training-Free Test-Time Adaptation that Improves Model Accuracy on ImageNet1K to 96%

DGX agent

arXiv:2604.04552v2 Announce Type: replace-cross Abstract: Ensemble methods are widely used to improve predictive performance, but their effectiveness often comes at the cost of increased memory usage

model-releasesarxiv-cs-ai
14 Apr 2026
Research

Towards Reasonable Concept Bottleneck Models

DGX agent

arXiv:2506.05014v2 Announce Type: replace-cross Abstract: We propose a novel, flexible, and efficient framework for designing Concept Bottleneck Models (CBMs) that enables practitioners to explicitly

researcharxiv-cs-ai
14 Apr 2026
Applications

Towards Situation-aware State Modeling for Air Traffic Flow Prediction

DGX agent

arXiv:2604.11198v1 Announce Type: new Abstract: Accurate air traffic prediction in the terminal airspace (TA) is pivotal for proactive air traffic management (ATM). However, existing data-driven appro

applicationsarxiv-cs-lg
14 Apr 2026
Research

Vibe-driven model-based engineering

DGX agent

arXiv:2604.10645v1 Announce Type: cross Abstract: There is a pressing need for better development methods and tools to keep up with the growing demand and increasing complexity of new software systems

researcharxiv-cs-ai
14 Apr 2026
Tutorials

What do your logits know? (The answer may surprise you!)

DGX agent

arXiv:2604.09885v1 Announce Type: new Abstract: Recent work has shown that probing model internals can reveal a wealth of information not apparent from the model generations. This poses the risk of un

tutorialsarxiv-cs-ai
14 Apr 2026
Model Releases

Dream to Fly: Model-Based Reinforcement Learning for Vision-Based Drone Flight

DGX agent

arXiv:2501.14377v2 Announce Type: replace Abstract: Autonomous drone racing has risen as a challenging robotic benchmark for testing the limits of learning, perception, planning, and control. Expert h

model-releasesarxiv-cs-ro
13 Apr 2026
Model Releases

DSVTLA: Deep Swin Vision Transformer-Based Transfer Learning Architecture for Multi-Type Cancer Histopathological Cancer Image Classification

DGX agent

arXiv:2604.09468v1 Announce Type: cross Abstract: In this study, we proposed a deep Swin-Vision Transformer-based transfer learning architecture for robust multi-cancer histopathological image classif

model-releasesarxiv-cs-cv
13 Apr 2026
Safety

Enhancing the Safety of Medical Vision-Language Models by Synthetic Demonstrations

DGX agent

arXiv:2506.09067v2 Announce Type: replace-cross Abstract: Generative medical vision-language models~(Med-VLMs) are primarily designed to generate complex textual information~(e.g., diagnostic reports)

safetyarxiv-cs-ai
13 Apr 2026
Agents

GeRM: A Generative Rendering Model From Physically Realistic to Photorealistic

DGX agent

arXiv:2604.09304v1 Announce Type: new Abstract: For decades, Physically-Based Rendering (PBR) is the fundation of synthesizing photorealisitic images, and therefore sometimes roughly referred as Photo

agentsarxiv-cs-cv
13 Apr 2026
Research

Grammar as a Behavioral Biometric: Using Cognitively Motivated Grammar Models for Authorship Verification

DGX agent

arXiv:2403.08462v3 Announce Type: replace Abstract: Authorship Verification (AV) is a key area of research in digital text forensics, which addresses the fundamental question of whether two texts were

researcharxiv-cs-cl
13 Apr 2026
Research

InsEdit: Towards Instruction-based Visual Editing via Data-Efficient Video Diffusion Models Adaptation

DGX agent

arXiv:2604.08646v1 Announce Type: new Abstract: Instruction-based video editing is a natural way to control video content with text, but adapting a video generation model into an editor usually appear

researcharxiv-cs-cv
13 Apr 2026
Agents

Koopman Operator Framework for Modeling and Control of Off-Road Vehicle on Deformable Terrain

DGX agent

arXiv:2603.28965v2 Announce Type: replace-cross Abstract: This work presents a hybrid physics-informed and data-driven modeling framework for predictive control of autonomous off-road vehicles operati

agentsarxiv-cs-ro
13 Apr 2026
Safety

Large Reasoning Models Learn Better Alignment from Flawed Thinking

DGX agent

arXiv:2510.00938v2 Announce Type: replace Abstract: Large reasoning models (LRMs) 'think' by generating structured chain-of-thought (CoT) before producing a final answer, yet they still lack the abili

safetyarxiv-cs-lg
13 Apr 2026
Model Releases

Parameterized Complexity Of Representing Models Of MSO Formulas

DGX agent

arXiv:2604.08707v1 Announce Type: new Abstract: Monadic second order logic (MSO2) plays an important role in parameterized complexity due to the Courcelle's theorem. This theorem states that the probl

model-releasesarxiv-cs-ai
13 Apr 2026
← Previous
1…168169170171172…1030
Next →