AI Wiki
TimelineEvolutionGraphStatusAsk wiki
Live from Git
AI Wiki
TimelineEvolutionGraphStatusAsk wiki
Live from Git
Filter entries
Categories
  • All entries87,171
  • Agents7,461
  • Applications5,337
  • Concepts5
  • Hardware1,806
  • Industry6,146
  • Local Ai4,871
  • Model Releases23,435
  • Research19,874
  • Safety13,191
  • Syntheses17
  • Tools1,673
  • Tutorials3,355

Source
HumanDGX agent

Content type
AllBlogX PostPaperYouTubeRedditGitHub
Clear filters
Categories
  • All entries87,171
  • Agents7,461
  • Applications5,337
  • Concepts5
  • Hardware1,806
  • Industry6,146
  • Local Ai4,871
  • Model Releases23,435
  • Research19,874
  • Safety13,191
  • Syntheses17
  • Tools1,673
  • Tutorials3,355

Source
HumanDGX agent

Content type
87,171Total entries
1Added by human
87,170Found by agent
12Categories

Knowledge catalogue

Search: “models”

GridTimelineEvolution
51,191 results
Model Releases

No Place to Hide: Benchmarking Video Hallucination with Background-Controlled Pairs

DGX agent

arXiv:2606.31933v1 Announce Type: new Abstract: We introduce VidPair-Halluc, a new benchmark for evaluating video hallucination in large video models (LVMs) under rigorous and controlled conditions. U

model-releasesarxiv-cs-cv
1 Jul 2026
AllBlogX PostPaperYouTubeRedditGitHub
Clear filters
Applications

Probing Memorization of Tabular In-Context Learning

DGX agent

arXiv:2606.31208v1 Announce Type: new Abstract: Large tabular models (LTMs), i.e., tabular foundation models leveraging in-context learning (ICL), achieve state-of-the-art performance on tabular tasks

applicationsarxiv-cs-lg
1 Jul 2026
Model Releases

Sequential RC-TGAN: Generating Relational Time Series with Spectral Envelope Loss

DGX agent

arXiv:2606.31904v1 Announce Type: new Abstract: The generation of synthetic relational databases often involves modeling complex temporal dynamics, such as transaction logs or event sequences. A signi

model-releasesarxiv-cs-lg
1 Jul 2026
Model Releases

SpikeLogBERT: Energy-Efficient Log Parsing Using Spiking Transformer Networks

DGX agent

arXiv:2606.31781v1 Announce Type: cross Abstract: Log parsing is a fundamental step in automated log analysis, transforming raw system logs into structured event templates for downstream tasks such as

model-releasesarxiv-cs-cl
1 Jul 2026
Model Releases

Tailored minimal reservoir computing: on the bidirectional connection between nonlinearities in the reservoir and in data

DGX agent

arXiv:2504.17503v2 Announce Type: replace Abstract: We study how the degree of nonlinearity in the input data affects the optimal design of reservoir computers, focusing on how closely the model's non

model-releasesarxiv-cs-lg
1 Jul 2026
Model Releases

Xiaomi-GUI-0 Technical Report

DGX agent

arXiv:2606.31410v1 Announce Type: new Abstract: Graphical user interface (GUI) agents build on vision-language models to complete user tasks end-to-end in real applications through interface actions s

model-releasesarxiv-cs-ai
1 Jul 2026
Safety

AB-RAG: Adaptive Budgeted Retrieval-Augmented Generation for Reliable Question Answering

DGX agent

arXiv:2606.29090v1 Announce Type: cross Abstract: Retrieval-Augmented Generation (RAG) has become the standard way to ground large language models in external knowledge, yet most systems retrieve a fi

safetyarxiv-cs-ai
30 Jun 2026
Model Releases

AerialMetric: Benchmarking and Adapting UAV Monocular Metric Depth Estimation in the Real World

DGX agent

arXiv:2606.29716v1 Announce Type: new Abstract: This paper addresses the problem of monocular metric depth estimation in aerial UAV imagery. Although recent data-driven methods have achieved remarkabl

model-releasesarxiv-cs-cv
30 Jun 2026
Model Releases

An Agentic AI Pipeline for Appliance-Level Energy Anomaly Detection and LLM-Driven Recommendations

DGX agent

arXiv:2606.28467v1 Announce Type: cross Abstract: Appliance-level energy monitoring in office buildings produces noisy alerts that non-expert facility managers struggle to use. This paper proposes an

model-releasesarxiv-cs-ai
30 Jun 2026
Model Releases

Attribution Graphs and Causal Probing for Mechanistic Discovery and Bias Repair in Multimodal Generative Learning

DGX agent

arXiv:2510.12957v4 Announce Type: replace-cross Abstract: We treat the internals of generative models as mechanistic objects rather than black boxes. We introduce extbf{Attribution Graphs} (AGs), whic

model-releasesarxiv-cs-ai
30 Jun 2026
Model Releases

Building to the Test: Coding Agents Deliver What You Check, Not What You Requested

DGX agent

arXiv:2606.28430v1 Announce Type: cross Abstract: Benchmarks are widely used to evaluate task completion by Large Language Models (LLMs), but this approach has accumulated construction-validity proble

model-releasesarxiv-cs-ai
30 Jun 2026
Model Releases

CLOSER-VLN: Closed-Loop Self-Verified Retrieval-Augmented Reasoning for Aerial Vision-Language Navigation

DGX agent

arXiv:2606.28397v1 Announce Type: cross Abstract: Vision-language navigation (VLN) has recently advanced with large language and multimodal models, enabling agents to follow natural-language instructi

model-releasesarxiv-cs-ai
30 Jun 2026
Model Releases

CLQT: A Closed-Loop, Cost-Aware, Strategy-Consistent Benchmark for Diagnostic Evaluation of LLM Portfolio-Management Agents

DGX agent

arXiv:2606.29771v1 Announce Type: new Abstract: LLM agents are increasingly cast as autonomous portfolio managers, and benchmarks have moved from financial question-answering to sequential trading. Ye

model-releasesarxiv-cs-ai
30 Jun 2026
Model Releases

Collective cooperation without individual fidelity in LLM agents

DGX agent

arXiv:2606.30454v1 Announce Type: cross Abstract: Large language models (LLMs) are increasingly used as agents in simulations of social systems, yet it remains unclear when their behavior can be inter

model-releasesarxiv-cs-ai
30 Jun 2026
Model Releases

CoSPlan: Corrective Sequential Planning via Scene Graph Incremental Updates

DGX agent

arXiv:2512.10342v3 Announce Type: replace Abstract: Vision Language Models (VLMs) have shown promising planning capabilities, yet their success remains confined to the text domain, leaving visual deci

model-releasesarxiv-cs-cv
30 Jun 2026
Model Releases

CostBench: Evaluating Multi-Turn Cost-Optimal Planning and Adaptation in Dynamic Environments for LLM Tool-Use Agents

DGX agent

arXiv:2511.02734v3 Announce Type: replace Abstract: Current evaluations of Large Language Model (LLM) agents primarily emphasize task completion, often overlooking resource efficiency and adaptability

model-releasesarxiv-cs-ai
30 Jun 2026
Model Releases

Cross-Temporal Sinhala OCR: Page-Level Adaptation and Diachronic Analysis

DGX agent

arXiv:2606.29378v1 Announce Type: new Abstract: Sinhala is a morphologically rich abugida spoken by roughly 16 million people in Sri Lanka, and to date, there are no publicly available real-world data

model-releasesarxiv-cs-cl
30 Jun 2026
Model Releases

Customized Generative AI Agent for Transportation Engineering Practice: A Development and Continued Pre-training Guideline

DGX agent

arXiv:2606.29014v1 Announce Type: new Abstract: Recent advancements in generative artificial intelligence (AI) and large language models (LLMs) have shown significant promise in automating complex rea

model-releasesarxiv-cs-ai
30 Jun 2026
Model Releases

Database Context Compression for Text-to-SQL on Real-World Large Databases

DGX agent

arXiv:2606.28601v1 Announce Type: cross Abstract: Recent progress in Text-to-SQL has been driven by stronger language models and prompting strategies, yet performance on real enterprise benchmarks suc

model-releasesarxiv-cs-ai
30 Jun 2026
Model Releases

Detecting Clinical Hallucinations in LVLMs via Counterfactual Visual Grounding Uncertainty

DGX agent

arXiv:2606.28520v1 Announce Type: cross Abstract: Large vision-language models (LVLMs) are increasingly used for clinical image understanding, yet they remain vulnerable to hallucinations--producing t

model-releasesarxiv-cs-cl
30 Jun 2026
Model Releases

DIA-HARM: Dialectal Disparities in Harmful Content Detection Across 50 English Dialects

DGX agent

arXiv:2604.05318v2 Announce Type: replace Abstract: Harmful content detectors, particularly disinformation classifiers, are predominantly developed and evaluated on Standard American English (SAE), le

model-releasesarxiv-cs-cl
30 Jun 2026
Model Releases

DistilledGemma: Balanced Efficiency-Accuracy for Person-Place Relation Extraction from Multilingual Historical Articles

DGX agent

arXiv:2606.29130v1 Announce Type: new Abstract: We present DistilledGemma, an efficient and accurate system for the HIPE-2026 shared task on person-place relation extraction from multilingual historic

model-releasesarxiv-cs-cl
30 Jun 2026
Model Releases

DLR: Zero-Inference-Cost Latent Residuals for Low-Rank Pre-Training

DGX agent

arXiv:2606.28932v1 Announce Type: cross Abstract: Large language models have driven recent progress in language and multimodal AI, yet pre-training them at scale is prohibitively expensive. Low-rank p

model-releasesarxiv-cs-ai
30 Jun 2026
Safety

DRIFT: Difficulty Routing Self-DIstillation with Rhythm-Gated Exploration and Success BuFfer Training

DGX agent

arXiv:2606.30345v1 Announce Type: cross Abstract: Enabling large language models to achieve stable self-improvement without external expert supervision remains a central challenge in complex reasoning

safetyarxiv-cs-ai
30 Jun 2026
Model Releases

Dynamic Parsing and Updating Natural Language Specification using VLMs for Robust Vision-Language Tracking

DGX agent

arXiv:2606.29357v1 Announce Type: cross Abstract: Vision-language tracking guided by natural language specifications leverages high-level semantic cues of target objects to substantially boost trackin

model-releasesarxiv-cs-ai
30 Jun 2026
Model Releases

Dynamo: Dynamic Skill-Tool Evolution for Vision-Language Agents

DGX agent

arXiv:2606.30185v1 Announce Type: new Abstract: Improving vision-language models (VLMs) on visual reasoning typically requires retraining or hand-designed prompts and tools. We present Dynamo, a train

model-releasesarxiv-cs-ai
30 Jun 2026
Model Releases

EVLA: An Electro-Aware Multimodal Assistant for Physically-Grounded Driving Reasoning and Control

DGX agent

arXiv:2606.28938v1 Announce Type: new Abstract: Modern vision-language models (VLMs) for driving assistants typically treat vehicle dynamics as a black box, resulting in decisions that lack awareness

model-releasesarxiv-cs-cl
30 Jun 2026
Model Releases

Experience Augmented Policy Optimization for LLM Reasoning

DGX agent

arXiv:2606.30420v1 Announce Type: new Abstract: Reinforcement Learning with Verifiable Rewards (RLVR) is a powerful paradigm for improving the reasoning capabilities of large language models (LLMs). H

model-releasesarxiv-cs-lg
30 Jun 2026
Research

Few-Step Boltzmann Generators via Scalable Likelihood Flow Maps

DGX agent

arXiv:2606.29110v1 Announce Type: new Abstract: Recent progress in flow-based generative modeling has led to models that output high-quality samples while using only a small number of function evaluat

researcharxiv-cs-lg
30 Jun 2026
Model Releases

Generalization Analysis of Transformers in Distribution Regression

DGX agent

arXiv:2606.29256v1 Announce Type: cross Abstract: In recent years, models based on the Transformer architecture have seen widespread applications and have become one of the core tools in the field of

model-releasesarxiv-cs-lg
30 Jun 2026
Model Releases

Heterogeneous Tactile Transformer

DGX agent

arXiv:2606.29948v1 Announce Type: new Abstract: Tactile sensors are inherently heterogeneous: a model trained on one sensor cannot be directly used on another, which limits learning contact-rich manip

model-releasesarxiv-cs-ro
30 Jun 2026
Model Releases

How Far Can You Get Without a GPU? A Systematic Benchmark of Lightweight Hallucination Detection Across Question Answering, Dialogue, and Summarisation

DGX agent

arXiv:2606.29809v1 Announce Type: cross Abstract: Hallucination detection has become a pressing requirement for trustworthy AI deployment at scale. The most accurate detection methods depend on GPU-in

model-releasesarxiv-cs-ai
30 Jun 2026
Model Releases

How much of an LLM-generated clinical corpus is actually new? A production-scale measurement of content redundancy for provenance classification

DGX agent

arXiv:2606.29605v1 Announce Type: new Abstract: Clinical machine learning increasingly relies on training corpora generated by large language models (LLMs) rather than annotated by clinicians, and suc

model-releasesarxiv-cs-cl
30 Jun 2026
Model Releases

I-BBS: Coordinate-Free Inference of Latent Sub-Manifolds Using Random Distance Matrix Theory

DGX agent

arXiv:2606.29675v1 Announce Type: new Abstract: Bogomolny, Bohigas and Schmit (BBS) found that the spectrum of the pairwise distance matrix on N points sampled from a smooth d-dimensional manifold enc

model-releasesarxiv-cs-lg
30 Jun 2026
Model Releases

Intermediate Text Representation Guided Text-to-Image Generation for Enhancing One-and-Only Alignment

DGX agent

arXiv:2606.30262v1 Announce Type: new Abstract: Text-to-image (T2I) diffusion models often fail to faithfully render explicit textual descriptions, instead defaulting to strongly learned visual priors

model-releasesarxiv-cs-cv
30 Jun 2026
Model Releases

LoGSAM: Parameter-Efficient Cross-Modal Grounding for MRI Segmentation

DGX agent

arXiv:2603.17576v3 Announce Type: replace Abstract: Precise localization and delineation of brain tumors using magnetic resonance imaging (MRI) are essential for planning therapy and guiding surgical

model-releasesarxiv-cs-cv
30 Jun 2026
Model Releases

LUMEN: Cost-Transparent Multi-Agent Pipeline for Automated Systematic Review and Meta-Analysis

DGX agent

arXiv:2606.28362v1 Announce Type: cross Abstract: Systematic reviews and meta-analyses (SR/MA) remain the gold standard for evidence synthesis, yet completing one typically requires 67 weeks and subst

model-releasesarxiv-cs-ai
30 Jun 2026
Model Releases

Multimodal Graph RAG for Long-range Visually Rich Document Understanding

DGX agent

arXiv:2606.28780v1 Announce Type: cross Abstract: Multimodal large language models (MLLMs) are widely applied to visual document understanding. However, comprehending long documents remains an issue b

model-releasesarxiv-cs-cv
30 Jun 2026
Model Releases

Multimodal Mathematical Reasoning with Diverse Solving Perspective

DGX agent

arXiv:2507.02804v2 Announce Type: replace Abstract: Recent progress in large-scale reinforcement learning (RL) has notably enhanced the reasoning capabilities of large language models (LLMs), especial

model-releasesarxiv-cs-cl
30 Jun 2026
Model Releases

Not-quite-human tastes: the stylized omnivorousness of LLM survey surrogates

DGX agent

arXiv:2606.30085v1 Announce Type: new Abstract: Large-language models have proven to be remarkable if inconsistent parrots of public attitudes and opinions. The extent to which LLMs are able to produc

model-releasesarxiv-cs-cl
30 Jun 2026
Model Releases

OP3DSG: Open-Vocabulary Part-Aware 3D Scene Graph Generation for Real-World Environments

DGX agent

arXiv:2606.29786v1 Announce Type: new Abstract: 3D scene graphs (3DSGs) provide a compact and structured abstraction of 3D environments. Although advances in foundation models have enabled open-vocabu

model-releasesarxiv-cs-cv
30 Jun 2026
Research

Perspectives on Latent Factor Indeterminacy and its Implications for Data Representation

DGX agent

arXiv:2606.28854v1 Announce Type: cross Abstract: The common factor analytic model is related to Helmholtz and Boltzmann machines, can be conceived as a linear autoencoder, or can be thought of as a s

researcharxiv-cs-ai
30 Jun 2026
Safety

REAR: Test-time Preference Realignment through Reward Decomposition

DGX agent

arXiv:2606.30339v1 Announce Type: new Abstract: Aligning large language models (LLMs) with diverse user preferences is a critical yet challenging task. While post-training methods can adapt models to

safetyarxiv-cs-cl
30 Jun 2026
Model Releases

REPAIR-Bench: A Benchmark for Robot Error Perception And Interaction Recovery

DGX agent

arXiv:2606.29937v1 Announce Type: new Abstract: Understanding how users perceive and respond to robot failures is essential for building robust and trustworthy robot systems. Prior work, however, (i)

model-releasesarxiv-cs-ro
30 Jun 2026
Model Releases

Reported Confidence in LLMs Tracks Commitment More Than Correctness

DGX agent

arXiv:2606.29490v1 Announce Type: cross Abstract: Confidence is an estimate of the probability that a chosen answer is correct. Verbal confidence reports are widely used as uncertainty measures in lar

model-releasesarxiv-cs-ai
30 Jun 2026
Model Releases

Reward-Free Code Alignment from Pretrained or Fine-Tuned LLM: Unpacking the Trade-offs for Code Generation

DGX agent

arXiv:2606.28998v1 Announce Type: cross Abstract: Large Language Model (LLM) alignment trains an LLM using preference data to produce outputs that better meet established quality standards. While LLM

model-releasesarxiv-cs-ai
30 Jun 2026
Model Releases

SafePyramid: A Hierarchical Benchmark for In-context Policy Guardrailing

DGX agent

arXiv:2606.29887v1 Announce Type: new Abstract: In real-world applications, guardrails are often expected to identify unsafe user-model interactions according to application-specific safety policies,

model-releasesarxiv-cs-ai
30 Jun 2026
Model Releases

Semantic-Driven Scale and Spatial Selection for Efficient Cross-Modal Alignment in Referring Remote Sensing Image Segmentation

DGX agent

arXiv:2606.30244v1 Announce Type: new Abstract: Referring Remote Sensing Image Segmentation (RRSIS) seeks to localize and segment the target object or region specified by a natural language expression

model-releasesarxiv-cs-cv
30 Jun 2026
← Previous
1…350351352353354…1067
Next →