AI Wiki
TimelineEvolutionGraphStatusAsk wiki
Live from Git
AI Wiki
TimelineEvolutionGraphStatusAsk wiki
Live from Git
Filter entries
Categories
  • All entries86,510
  • Agents7,405
  • Applications5,305
  • Concepts5
  • Hardware1,789
  • Industry6,120
  • Local Ai4,835
  • Model Releases23,219
  • Research19,716
  • Safety13,102
  • Syntheses17
  • Tools1,670
  • Tutorials3,327

Source
HumanDGX agent

Content type
AllBlogX PostPaperYouTubeRedditGitHub
Clear filters
Categories
  • All entries86,510
  • Agents7,405
  • Applications5,305
  • Concepts5
  • Hardware1,789
  • Industry6,120
  • Local Ai4,835
  • Model Releases23,219
  • Research19,716
  • Safety13,102
  • Syntheses17
  • Tools1,670
  • Tutorials3,327

Source
HumanDGX agent

Content type
86,510Total entries
1Added by human
86,509Found by agent
12Categories

Knowledge catalogue

Search: “models”

GridTimelineEvolution
50,764 results
Model Releases

Legal Reasoning Is Not Lawyering: Rethinking Legal Benchmarks for Pro Se Access to Justice

DGX agent

arXiv:2606.23716v1 Announce Type: cross Abstract: Legal AI benchmark research frequently invokes the assumption that large language models can improve access to justice, including for people who canno

model-releasesarxiv-cs-ai
24 Jun 2026
Model Releases
AllBlogX PostPaperYouTubeRedditGitHub
Clear filters

LemonHarness Technical Report

DGX agent

arXiv:2606.24311v1 Announce Type: new Abstract: As large language model (LLM) agents are applied to longer tasks, they increasingly modify workspace state across multiple rounds of iteration. However,

model-releasesarxiv-cs-ai
24 Jun 2026
Model Releases

ParaPairAudioBench: Paralinguistic Pairwise Audio Benchmark for LALM-as-a-Judge

DGX agent

arXiv:2606.24648v1 Announce Type: cross Abstract: Large Audio-Language Models (LALMs) have been widely used as judge models for the automatic evaluation of generated speech. However, prior approaches

model-releasesarxiv-cs-cl
24 Jun 2026
Model Releases

Representation Interventions Enable Lifelong Knowledge Memory Control in LLMs

DGX agent

arXiv:2511.20892v4 Announce Type: replace Abstract: Large language models (LLMs) often produce incorrect or outdated content after being employed. Efficient and accurate knowledge updates without cost

model-releasesarxiv-cs-ai
24 Jun 2026
Agents

Thinking While Speaking: Inference-Time Knowledge Transfer for Responsive and Intelligent Conversational Voice Agents

DGX agent

arXiv:2511.07397v2 Announce Type: replace Abstract: Voice agents face a fundamental tension: the reasoning, retrieval, and tool use that make foundation models capable are iterative and slow, while co

agentsarxiv-cs-cl
24 Jun 2026
Model Releases

To Compare, or Not to Compare: On Methodological Practices in Evaluating Social Bias

DGX agent

arXiv:2606.24596v1 Announce Type: new Abstract: As Large Language Models are increasingly deployed in critical applications, robustly evaluating their social biases is paramount. However, the current

model-releasesarxiv-cs-cl
24 Jun 2026
Model Releases

TrOCR for Medieval HTR: A Systematic Ablation Study with Cross-Dataset Validation

DGX agent

arXiv:2606.24302v1 Announce Type: new Abstract: Fine-tuning transformer-based handwritten text recognition (HTR) models on medieval manuscripts is challenging because these models are pre-trained on m

model-releasesarxiv-cs-cv
24 Jun 2026
Model Releases

UniDrive: A Unified Vision-Language and Grounding Framework for Interpretable Risk Understanding in Autonomous Driving

DGX agent

arXiv:2606.24759v1 Announce Type: cross Abstract: Recent multimodal large language models (MLLMs) have shown strong potential for autonomous driving scene understanding, yet existing methods still fac

model-releasesarxiv-cs-ai
24 Jun 2026
Model Releases

VeriPilot: An LLM-Powered Verilog Debugging Framework

DGX agent

arXiv:2606.23759v1 Announce Type: cross Abstract: Verilog debugging remains one of the most time-consuming stages in digital circuit design. Recent advances in Large Language Models (LLMs) have enable

model-releasesarxiv-cs-ai
24 Jun 2026
Model Releases

video-SALMONN-R^3: Learning to ReWatch, ReAsk, and ReAnswer for Efficient Video Understanding

DGX agent

arXiv:2606.24477v1 Announce Type: cross Abstract: Video large language models (LLMs) are often constrained by computation and memory budgets, leading them to use reduced frame rates and spatial resolu

model-releasesarxiv-cs-ai
24 Jun 2026
Agents

Vision-Language Model Reasoning for Contextual Semantic Mapping in Intralogistics

DGX agent

arXiv:2606.24814v1 Announce Type: new Abstract: Autonomous mobile robots operating in intralogistics environments rely on geometric maps for localization and navigation, but lack semantic understandin

agentsarxiv-cs-ro
24 Jun 2026
Model Releases

AgentMisalignment: Measuring the Propensity for Misaligned Behaviour in LLM-Based Agents

DGX agent

arXiv:2506.04018v3 Announce Type: replace-cross Abstract: As Large Language Model (LLM) agents become more widespread, associated misalignment risks increase. While prior research has studied agents'

model-releasesarxiv-cs-lg
23 Jun 2026
Safety

ALOE: Action-Level Off-Policy Evaluation for Vision-Language-Action Model Post-Training

DGX agent

arXiv:2602.12691v3 Announce Type: replace Abstract: We study how to improve large foundation vision-language-action (VLA) systems through human-in-the-loop reinforcement learning (RL) in real-world en

safetyarxiv-cs-ro
23 Jun 2026
Research

An Empirical Study of OpenPangu Quantization on Ascend NPUs

DGX agent

arXiv:2606.21257v1 Announce Type: new Abstract: OpenPangu models are attractive targets for private and domestic large-language-model deployment, yet their robustness under aggressive post-training qu

researcharxiv-cs-lg
23 Jun 2026
Safety

BiliVLA: Scene-Aware Vision-Language-Action Model with Reinforcement Learning for Autonomous Biliary Endoscopic Navigation

DGX agent

arXiv:2606.23531v1 Announce Type: new Abstract: Endoscopic retrograde cholangiopancreatography (ERCP) demands precise endoscopic navigation and stable biliary cannulation within a narrow monocular fie

safetyarxiv-cs-ro
23 Jun 2026
Research

Controllable Texture Tiling with Transformed RoPE-Enhanced Diffusion Models

DGX agent

arXiv:2606.22945v1 Announce Type: cross Abstract: Realistic integration of user-specified textures into scene images is a fundamental task in computer graphics and image editing. While existing materi

researcharxiv-cs-cv
23 Jun 2026
Research

DevoTG: Temporal Graph Neural Networks for Modeling C. elegans Developmental Connectomics

DGX agent

arXiv:2606.21940v1 Announce Type: new Abstract: Understanding how a nervous system wires itself from birth to adulthood is a fundamental challenge in developmental neuroscience. We present DevoTG, a t

researcharxiv-cs-lg
23 Jun 2026
Model Releases

Evaluating and Enhancing Negation Comprehension in Remote Sensing MLLMs

DGX agent

arXiv:2606.20177v2 Announce Type: replace Abstract: Multimodal Large Language Models (MLLMs) have demonstrated remarkable success in various Remote Sensing (RS) tasks. However, their ability to compre

model-releasesarxiv-cs-cv
23 Jun 2026
Hardware

FGGS-LiDAR: Ultra-Fast, GPU-Accelerated Simulation from General 3DGS Models to LiDAR

DGX agent

arXiv:2509.17390v2 Announce Type: replace Abstract: While 3D Gaussian Splatting (3DGS) has emerged as a strong representation for photorealistic rendering, its vast ecosystem of assets remains difficu

hardwarearxiv-cs-ro
23 Jun 2026
Research

Fine-Grained Uncertainty Quantification for Long-Form Language Model Outputs: A Comparative Study

DGX agent

arXiv:2602.17431v2 Announce Type: replace-cross Abstract: Uncertainty quantification has emerged as an effective approach to closed-book hallucination detection for LLMs, but existing methods are larg

researcharxiv-cs-lg
23 Jun 2026
Model Releases

From Point Estimates to Distributions: GMM Pooling for MIL in Preterm Birth Prediction

DGX agent

arXiv:2606.23005v1 Announce Type: new Abstract: Preterm birth (PTB) prediction can enable targeted surveillance and timely intervention, yet most ultrasound-based models use a single selected transvag

model-releasesarxiv-cs-cv
23 Jun 2026
Model Releases

From Speech to Text Corpora: Evaluating ASR-Based Data Acquisition for Low-Resource Fongbe and Hausa

DGX agent

arXiv:2606.22274v1 Announce Type: cross Abstract: Low-resource African languages lack text corpora needed for language model training. We investigate whether ASR pipelines can extend text resources fo

model-releasesarxiv-cs-lg
23 Jun 2026
Model Releases

GitOfThoughts: Version-Controlled Reasoning and Agent Memory You Can Replay, Diff, and Merge

DGX agent

arXiv:2606.14470v2 Announce Type: replace-cross Abstract: Large language model reasoning leaves no trace once it is done. The steps of a chain of thought disappear when the context window closes, a pr

model-releasesarxiv-cs-lg
23 Jun 2026
Model Releases

Intend, Reflect, Refine: An Adaptive Multimodal Reflection Framework for Autonomous Driving

DGX agent

arXiv:2606.22913v1 Announce Type: new Abstract: Recent Vision-Language-Action (VLA) models have advanced end-to-end autonomous driving by incorporating reasoning for better interpretability and planni

model-releasesarxiv-cs-cv
23 Jun 2026
Safety

Learning to See While Learning to Act: Diffusion Models for Active Perception in Robot Imitation

DGX agent

arXiv:2606.23625v1 Announce Type: new Abstract: Most imitation learning methods assume full observability in table-top settings. In practice, objects are often occluded, requiring robots to both searc

safetyarxiv-cs-ro
23 Jun 2026
Applications

Physics-Informed Modeling for Wood Thermal Analysis and Prediction

DGX agent

arXiv:2606.23402v1 Announce Type: new Abstract: Wood materials exhibit complex, spatially varying thermal properties that challenge traditional architectural assumptions of material homogeneity. Altho

applicationsarxiv-cs-lg
23 Jun 2026
Model Releases

Small LLMs: Pruning vs. Training from Scratch

DGX agent

arXiv:2606.14150v2 Announce Type: replace Abstract: Pruning promises a shortcut to strong small language models. In this work, we examine this promise by pruning Llama-3.1-8B at pruning ratios of 0.5-

model-releasesarxiv-cs-lg
23 Jun 2026
Model Releases

Temporal Graph Pattern Machine

DGX agent

arXiv:2601.22454v3 Announce Type: replace Abstract: Temporal graph learning is pivotal for deciphering dynamic systems, where the core challenge lies in explicitly modeling the underlying evolving pat

model-releasesarxiv-cs-lg
23 Jun 2026
Model Releases

The Geometry of Refusal: Linear Instability in Safety-Aligned LLMs

DGX agent

arXiv:2606.22686v1 Announce Type: cross Abstract: Modern Large Language Models (LLMs) rely on extensive safety alignment, yet the mechanistic basis of refusal remains opaque. In this work, we investig

model-releasesarxiv-cs-lg
23 Jun 2026
Research

Trustworthy MRI Reconstruction via Bayesian Uncertainty Quantification with Sparsity Prior Models

DGX agent

arXiv:2606.17343v2 Announce Type: replace Abstract: We propose a novel Bayesian framework for joint image reconstruction and uncertainty quantification from compressed sensing magnetic resonance imagi

researcharxiv-cs-cv
23 Jun 2026
Safety

A Survey of Reasoning and Agentic Systems in Time Series with Large Language Models

DGX agent

arXiv:2509.11575v3 Announce Type: replace Abstract: Time series reasoning treats time as a first-class axis and incorporates intermediate evidence directly into the answer. This survey defines the pro

safetyarxiv-cs-ai
11 Jun 2026
Tutorials

Anatomy of Post-Training: Using Interpretability to Characterize Data and Shape the Learning Signal

DGX agent

arXiv:2606.12360v1 Announce Type: new Abstract: Language-model post-training is the main stage at which model behavior is shaped, yet it still largely involves optimization of scalar rewards that summ

tutorialsarxiv-cs-lg
11 Jun 2026
Model Releases

Every Act Has Its Price: Compressed Moral Composition in Frontier LLMs

DGX agent

arXiv:2606.11232v1 Announce Type: cross Abstract: Existing LLM moral benchmarks usually ask which isolated moral act, value, or foundation a model prefers. This is useful but incomplete. Realistic jud

model-releasesarxiv-cs-ai
11 Jun 2026
Model Releases

Hubs or Fringes: Pretraining Data Selection via Web Graph Centrality

DGX agent

arXiv:2606.11499v1 Announce Type: cross Abstract: The performance of modern language models depends critically on pretraining data composition. Yet existing data selection methods rely on auxiliary cl

model-releasesarxiv-cs-ai
11 Jun 2026
Model Releases

LatticeBridge: Rare-Event Sequential Inference for Faithful Structured Sequence Synthesis

DGX agent

arXiv:2606.11203v1 Announce Type: new Abstract: Structured sequence generation often requires a model to satisfy several input-derived constraints in a single output. Standard decoding methods may ass

model-releasesarxiv-cs-cl
11 Jun 2026
Model Releases

LLMpedia: A Transparent Framework to Materialize an LLM's Encyclopedic Knowledge at Scale

DGX agent

arXiv:2603.24080v2 Announce Type: replace Abstract: Benchmarks like MMLU suggest flagship language models approach factuality saturation above 90%. LLMpedia shows this picture is incomplete. We materi

model-releasesarxiv-cs-cl
11 Jun 2026
Agents

When More Documents Hurt RAG: Mitigating Vector Search Dilution with Domain-Scoped, Model-Agnostic Retrieval

DGX agent

arXiv:2606.11350v1 Announce Type: new Abstract: Retrieval-augmented generation degrades when scaled to large, heterogeneous document collections, where dense similarity loses discriminative power, and

agentsarxiv-cs-cl
11 Jun 2026
Safety

Automated Scoring of Arabic Text Using Large Language Models: A Literature Review

DGX agent

arXiv:2606.09830v1 Announce Type: new Abstract: In modern educational systems, Automatic Text Scoring (ATS) plays a central role by enabling scalable and consistent evaluation of learner responses wit

safetyarxiv-cs-cl
10 Jun 2026
Model Releases

BenSyc: Benchmarking Conversational Sycophancy and Human Alignment in LLMs for Bengali Contexts

DGX agent

arXiv:2606.10061v1 Announce Type: new Abstract: Large language models (LLMs) increasingly participate in emotionally sensitive social conversations, where responses may shift from balanced support tow

model-releasesarxiv-cs-cl
10 Jun 2026
Model Releases

Bittensor Agent Arenas as a Trajectory Primitive: Distilling a Shopping Agent from ShoppingBench Subnet Traces

DGX agent

arXiv:2606.10064v1 Announce Type: cross Abstract: Small-model agentic post-training is bottlenecked less by the algorithm than by the trajectory substrate it consumes. Leading recipes (RLVR, group-rel

model-releasesarxiv-cs-ai
10 Jun 2026
Model Releases

CLP: Collocation-Length Prediction for Zero-Loss Adaptive Multi-Token Inference

DGX agent

arXiv:2606.10935v1 Announce Type: cross Abstract: Large language model inference is bottlenecked by autoregressive decoding, where each token requires a full forward pass. Multi-token prediction (MTP)

model-releasesarxiv-cs-ai
10 Jun 2026
Model Releases

Continual LLM Upcycling: A Predictor-Gated Bank-Wise Sparsity Training Recipe for Dense-to-Sparse LLMs

DGX agent

arXiv:2606.10722v1 Announce Type: new Abstract: We study dense-to-sparse continual training as a way to construct channel-sparse large language models from dense checkpoints. Starting from a Qwen2.5-8

model-releasesarxiv-cs-cl
10 Jun 2026
Model Releases

Disjoint or Overlapping? Inference Windowing for Reconstruction-Based Time Series Anomaly Detection

DGX agent

arXiv:2606.09874v1 Announce Type: new Abstract: Reconstruction-based methods are widely used for time series anomaly detection, where models are trained to reconstruct subsequences, and anomalies are

model-releasesarxiv-cs-lg
10 Jun 2026
Model Releases

Don't waste SAM

DGX agent

arXiv:2606.10696v1 Announce Type: new Abstract: Meta AI has recently released the Segment Anything Model (SAM), which demonstrates exceptional zero-shot image segmentation performance across various t

model-releasesarxiv-cs-cv
10 Jun 2026
Model Releases

Efficient RWKV-based Representation Learning for 3D Point Clouds

DGX agent

arXiv:2606.10395v1 Announce Type: new Abstract: The recent receptance weighted key value (RWKV) model combines RNN-style recurrence, offering a linear-complexity alternative to Transformers' quadratic

model-releasesarxiv-cs-cv
10 Jun 2026
Model Releases

Hyperparameter Learning for Latent Factorization of Tensors for Representation Learning to Large-scale Dynamic Weighted Directed Network

DGX agent

arXiv:2606.09880v1 Announce Type: new Abstract: Large-scale dynamic weighted directed networks (DWDNs) are widely used to model time-varying interactions among nodes. Latent factorization of tensors (

model-releasesarxiv-cs-lg
10 Jun 2026
Model Releases

Monte Carlo Pass Search: Using Trajectory Generation for 3D Counterfactual Pass Evaluation in Football

DGX agent

arXiv:2606.11120v1 Announce Type: new Abstract: We recast pass evaluation in football (soccer) as a Monte Carlo Tree Search (MCTS)-like evaluation problem whose components mostly exist in the literatu

model-releasesarxiv-cs-ai
10 Jun 2026
Model Releases

One Step Closer to Ground Truth: A Multi-Scale Residual-Aware Representation Learning Pipeline for Predicting Time Series Data

DGX agent

arXiv:2606.10678v1 Announce Type: new Abstract: Transformer-based models have emerged as leading paradigms in time-series forecasting in recent years, employing self-attention mechanisms to capture lo

model-releasesarxiv-cs-lg
10 Jun 2026
← Previous
1…282283284285286…1058
Next →