AI Wiki
TimelineEvolutionGraphStatusAsk wiki
Live from Git
AI Wiki
TimelineEvolutionGraphStatusAsk wiki
Live from Git
Filter entries
Categories
  • All entries86,457
  • Agents7,399
  • Applications5,302
  • Concepts5
  • Hardware1,786
  • Industry6,117
  • Local Ai4,835
  • Model Releases23,193
  • Research19,715
  • Safety13,094
  • Syntheses17
  • Tools1,670
  • Tutorials3,324

Source
HumanDGX agent

Content type
AllBlogX PostPaperYouTubeRedditGitHub
Clear filters
Categories
  • All entries86,457
  • Agents7,399
  • Applications5,302
  • Concepts5
  • Hardware1,786
  • Industry6,117
  • Local Ai4,835
  • Model Releases23,193
  • Research19,715
  • Safety13,094
  • Syntheses17
  • Tools1,670
  • Tutorials3,324

Source
HumanDGX agent

86,457Total entries
1Added by human
86,456Found by agent
12Categories

Knowledge catalogue

Search: “models”

GridTimelineEvolution
62,039 results
24 Jun 2026

LemonHarness Technical Report

Model ReleasesDGX agent

arXiv:2606.24311v1 Announce Type: new Abstract: As large language model (LLM) agents are applied to longer tasks, they increasingly modify workspace state across multiple rounds of iteration. However,

ParaPairAudioBench: Paralinguistic Pairwise Audio Benchmark for LALM-as-a-Judge

Model ReleasesDGX agent

arXiv:2606.24648v1 Announce Type: cross Abstract: Large Audio-Language Models (LALMs) have been widely used as judge models for the automatic evaluation of generated speech. However, prior approaches

Representation Interventions Enable Lifelong Knowledge Memory Control in LLMs

Model ReleasesDGX agent

arXiv:2511.20892v4 Announce Type: replace Abstract: Large language models (LLMs) often produce incorrect or outdated content after being employed. Efficient and accurate knowledge updates without cost

Content type
AllBlogX PostPaperYouTubeRedditGitHub
Clear filters

Thinking While Speaking: Inference-Time Knowledge Transfer for Responsive and Intelligent Conversational Voice Agents

AgentsDGX agent

arXiv:2511.07397v2 Announce Type: replace Abstract: Voice agents face a fundamental tension: the reasoning, retrieval, and tool use that make foundation models capable are iterative and slow, while co

To Compare, or Not to Compare: On Methodological Practices in Evaluating Social Bias

Model ReleasesDGX agent

arXiv:2606.24596v1 Announce Type: new Abstract: As Large Language Models are increasingly deployed in critical applications, robustly evaluating their social biases is paramount. However, the current

TrOCR for Medieval HTR: A Systematic Ablation Study with Cross-Dataset Validation

Model ReleasesDGX agent

arXiv:2606.24302v1 Announce Type: new Abstract: Fine-tuning transformer-based handwritten text recognition (HTR) models on medieval manuscripts is challenging because these models are pre-trained on m

UniDrive: A Unified Vision-Language and Grounding Framework for Interpretable Risk Understanding in Autonomous Driving

Model ReleasesDGX agent

arXiv:2606.24759v1 Announce Type: cross Abstract: Recent multimodal large language models (MLLMs) have shown strong potential for autonomous driving scene understanding, yet existing methods still fac

VeriPilot: An LLM-Powered Verilog Debugging Framework

Model ReleasesDGX agent

arXiv:2606.23759v1 Announce Type: cross Abstract: Verilog debugging remains one of the most time-consuming stages in digital circuit design. Recent advances in Large Language Models (LLMs) have enable

video-SALMONN-R^3: Learning to ReWatch, ReAsk, and ReAnswer for Efficient Video Understanding

Model ReleasesDGX agent

arXiv:2606.24477v1 Announce Type: cross Abstract: Video large language models (LLMs) are often constrained by computation and memory budgets, leading them to use reduced frame rates and spatial resolu

Vision-Language Model Reasoning for Contextual Semantic Mapping in Intralogistics

AgentsDGX agent

arXiv:2606.24814v1 Announce Type: new Abstract: Autonomous mobile robots operating in intralogistics environments rely on geometric maps for localization and navigation, but lack semantic understandin

23 Jun 2026

AgentMisalignment: Measuring the Propensity for Misaligned Behaviour in LLM-Based Agents

Model ReleasesDGX agent

arXiv:2506.04018v3 Announce Type: replace-cross Abstract: As Large Language Model (LLM) agents become more widespread, associated misalignment risks increase. While prior research has studied agents'

ALOE: Action-Level Off-Policy Evaluation for Vision-Language-Action Model Post-Training

SafetyDGX agent

arXiv:2602.12691v3 Announce Type: replace Abstract: We study how to improve large foundation vision-language-action (VLA) systems through human-in-the-loop reinforcement learning (RL) in real-world en

An Empirical Study of OpenPangu Quantization on Ascend NPUs

ResearchDGX agent

arXiv:2606.21257v1 Announce Type: new Abstract: OpenPangu models are attractive targets for private and domestic large-language-model deployment, yet their robustness under aggressive post-training qu

BiliVLA: Scene-Aware Vision-Language-Action Model with Reinforcement Learning for Autonomous Biliary Endoscopic Navigation

SafetyDGX agent

arXiv:2606.23531v1 Announce Type: new Abstract: Endoscopic retrograde cholangiopancreatography (ERCP) demands precise endoscopic navigation and stable biliary cannulation within a narrow monocular fie

Controllable Texture Tiling with Transformed RoPE-Enhanced Diffusion Models

ResearchDGX agent

arXiv:2606.22945v1 Announce Type: cross Abstract: Realistic integration of user-specified textures into scene images is a fundamental task in computer graphics and image editing. While existing materi

DevoTG: Temporal Graph Neural Networks for Modeling C. elegans Developmental Connectomics

ResearchDGX agent

arXiv:2606.21940v1 Announce Type: new Abstract: Understanding how a nervous system wires itself from birth to adulthood is a fundamental challenge in developmental neuroscience. We present DevoTG, a t

Evaluating and Enhancing Negation Comprehension in Remote Sensing MLLMs

Model ReleasesDGX agent

arXiv:2606.20177v2 Announce Type: replace Abstract: Multimodal Large Language Models (MLLMs) have demonstrated remarkable success in various Remote Sensing (RS) tasks. However, their ability to compre

FGGS-LiDAR: Ultra-Fast, GPU-Accelerated Simulation from General 3DGS Models to LiDAR

HardwareDGX agent

arXiv:2509.17390v2 Announce Type: replace Abstract: While 3D Gaussian Splatting (3DGS) has emerged as a strong representation for photorealistic rendering, its vast ecosystem of assets remains difficu

Fine-Grained Uncertainty Quantification for Long-Form Language Model Outputs: A Comparative Study

ResearchDGX agent

arXiv:2602.17431v2 Announce Type: replace-cross Abstract: Uncertainty quantification has emerged as an effective approach to closed-book hallucination detection for LLMs, but existing methods are larg

From Point Estimates to Distributions: GMM Pooling for MIL in Preterm Birth Prediction

Model ReleasesDGX agent

arXiv:2606.23005v1 Announce Type: new Abstract: Preterm birth (PTB) prediction can enable targeted surveillance and timely intervention, yet most ultrasound-based models use a single selected transvag

From Speech to Text Corpora: Evaluating ASR-Based Data Acquisition for Low-Resource Fongbe and Hausa

Model ReleasesDGX agent

arXiv:2606.22274v1 Announce Type: cross Abstract: Low-resource African languages lack text corpora needed for language model training. We investigate whether ASR pipelines can extend text resources fo

GitOfThoughts: Version-Controlled Reasoning and Agent Memory You Can Replay, Diff, and Merge

Model ReleasesDGX agent

arXiv:2606.14470v2 Announce Type: replace-cross Abstract: Large language model reasoning leaves no trace once it is done. The steps of a chain of thought disappear when the context window closes, a pr

Intend, Reflect, Refine: An Adaptive Multimodal Reflection Framework for Autonomous Driving

Model ReleasesDGX agent

arXiv:2606.22913v1 Announce Type: new Abstract: Recent Vision-Language-Action (VLA) models have advanced end-to-end autonomous driving by incorporating reasoning for better interpretability and planni

Learning to See While Learning to Act: Diffusion Models for Active Perception in Robot Imitation

SafetyDGX agent

arXiv:2606.23625v1 Announce Type: new Abstract: Most imitation learning methods assume full observability in table-top settings. In practice, objects are often occluded, requiring robots to both searc

Physics-Informed Modeling for Wood Thermal Analysis and Prediction

ApplicationsDGX agent

arXiv:2606.23402v1 Announce Type: new Abstract: Wood materials exhibit complex, spatially varying thermal properties that challenge traditional architectural assumptions of material homogeneity. Altho

Shared infrastructure, isolated tenants: Pool model multi-tenancy with Amazon Bedrock AgentCore

TutorialsDGX agent

In this post, you will learn patterns for implementing production-ready multi-tenant systems using Amazon Bedrock AgentCore. You will see these patterns demonstrated through healthcare AI agents that

Small LLMs: Pruning vs. Training from Scratch

Model ReleasesDGX agent

arXiv:2606.14150v2 Announce Type: replace Abstract: Pruning promises a shortcut to strong small language models. In this work, we examine this promise by pruning Llama-3.1-8B at pruning ratios of 0.5-

Temporal Graph Pattern Machine

Model ReleasesDGX agent

arXiv:2601.22454v3 Announce Type: replace Abstract: Temporal graph learning is pivotal for deciphering dynamic systems, where the core challenge lies in explicitly modeling the underlying evolving pat

The Geometry of Refusal: Linear Instability in Safety-Aligned LLMs

Model ReleasesDGX agent

arXiv:2606.22686v1 Announce Type: cross Abstract: Modern Large Language Models (LLMs) rely on extensive safety alignment, yet the mechanistic basis of refusal remains opaque. In this work, we investig

Trustworthy MRI Reconstruction via Bayesian Uncertainty Quantification with Sparsity Prior Models

ResearchDGX agent

arXiv:2606.17343v2 Announce Type: replace Abstract: We propose a novel Bayesian framework for joint image reconstruction and uncertainty quantification from compressed sensing magnetic resonance imagi

21 Jun 2026

Ideogram making 2 horrible precedent and we need to oppose that. BF16 weights not published and ridiculous model embedded censorship

Local AiDGX agent

I cannot provide a factual summary for a knowledge base without accessing the actual content. Reddit post titles often contain subjective language and don't reliably convey the full argument. To creat

11 Jun 2026

A Survey of Reasoning and Agentic Systems in Time Series with Large Language Models

SafetyDGX agent

arXiv:2509.11575v3 Announce Type: replace Abstract: Time series reasoning treats time as a first-class axis and incorporates intermediate evidence directly into the answer. This survey defines the pro

Anatomy of Post-Training: Using Interpretability to Characterize Data and Shape the Learning Signal

TutorialsDGX agent

arXiv:2606.12360v1 Announce Type: new Abstract: Language-model post-training is the main stage at which model behavior is shaped, yet it still largely involves optimization of scalar rewards that summ

Every Act Has Its Price: Compressed Moral Composition in Frontier LLMs

Model ReleasesDGX agent

arXiv:2606.11232v1 Announce Type: cross Abstract: Existing LLM moral benchmarks usually ask which isolated moral act, value, or foundation a model prefers. This is useful but incomplete. Realistic jud

Hubs or Fringes: Pretraining Data Selection via Web Graph Centrality

Model ReleasesDGX agent

arXiv:2606.11499v1 Announce Type: cross Abstract: The performance of modern language models depends critically on pretraining data composition. Yet existing data selection methods rely on auxiliary cl

LatticeBridge: Rare-Event Sequential Inference for Faithful Structured Sequence Synthesis

Model ReleasesDGX agent

arXiv:2606.11203v1 Announce Type: new Abstract: Structured sequence generation often requires a model to satisfy several input-derived constraints in a single output. Standard decoding methods may ass

LLMpedia: A Transparent Framework to Materialize an LLM's Encyclopedic Knowledge at Scale

Model ReleasesDGX agent

arXiv:2603.24080v2 Announce Type: replace Abstract: Benchmarks like MMLU suggest flagship language models approach factuality saturation above 90%. LLMpedia shows this picture is incomplete. We materi

When More Documents Hurt RAG: Mitigating Vector Search Dilution with Domain-Scoped, Model-Agnostic Retrieval

AgentsDGX agent

arXiv:2606.11350v1 Announce Type: new Abstract: Retrieval-augmented generation degrades when scaled to large, heterogeneous document collections, where dense similarity loses discriminative power, and

10 Jun 2026

Automated Scoring of Arabic Text Using Large Language Models: A Literature Review

SafetyDGX agent

arXiv:2606.09830v1 Announce Type: new Abstract: In modern educational systems, Automatic Text Scoring (ATS) plays a central role by enabling scalable and consistent evaluation of learner responses wit

BenSyc: Benchmarking Conversational Sycophancy and Human Alignment in LLMs for Bengali Contexts

Model ReleasesDGX agent

arXiv:2606.10061v1 Announce Type: new Abstract: Large language models (LLMs) increasingly participate in emotionally sensitive social conversations, where responses may shift from balanced support tow

Bittensor Agent Arenas as a Trajectory Primitive: Distilling a Shopping Agent from ShoppingBench Subnet Traces

Model ReleasesDGX agent

arXiv:2606.10064v1 Announce Type: cross Abstract: Small-model agentic post-training is bottlenecked less by the algorithm than by the trajectory substrate it consumes. Leading recipes (RLVR, group-rel

Claude Fable won’t answer basic biology questions

Model ReleasesDGX agent

Anthropic just released Claude Fable 5, calling it the most powerful AI model it has ever made widely available and praising its skills in biology, among others. But the model won't answer basic biolo

CLP: Collocation-Length Prediction for Zero-Loss Adaptive Multi-Token Inference

Model ReleasesDGX agent

arXiv:2606.10935v1 Announce Type: cross Abstract: Large language model inference is bottlenecked by autoregressive decoding, where each token requires a full forward pass. Multi-token prediction (MTP)

Continual LLM Upcycling: A Predictor-Gated Bank-Wise Sparsity Training Recipe for Dense-to-Sparse LLMs

Model ReleasesDGX agent

arXiv:2606.10722v1 Announce Type: new Abstract: We study dense-to-sparse continual training as a way to construct channel-sparse large language models from dense checkpoints. Starting from a Qwen2.5-8

Disjoint or Overlapping? Inference Windowing for Reconstruction-Based Time Series Anomaly Detection

Model ReleasesDGX agent

arXiv:2606.09874v1 Announce Type: new Abstract: Reconstruction-based methods are widely used for time series anomaly detection, where models are trained to reconstruct subsequences, and anomalies are

Don't waste SAM

Model ReleasesDGX agent

arXiv:2606.10696v1 Announce Type: new Abstract: Meta AI has recently released the Segment Anything Model (SAM), which demonstrates exceptional zero-shot image segmentation performance across various t

Efficient RWKV-based Representation Learning for 3D Point Clouds

Model ReleasesDGX agent

arXiv:2606.10395v1 Announce Type: new Abstract: The recent receptance weighted key value (RWKV) model combines RNN-style recurrence, offering a linear-complexity alternative to Transformers' quadratic

Hyperparameter Learning for Latent Factorization of Tensors for Representation Learning to Large-scale Dynamic Weighted Directed Network

Model ReleasesDGX agent

arXiv:2606.09880v1 Announce Type: new Abstract: Large-scale dynamic weighted directed networks (DWDNs) are widely used to model time-varying interactions among nodes. Latent factorization of tensors (

Monte Carlo Pass Search: Using Trajectory Generation for 3D Counterfactual Pass Evaluation in Football

Model ReleasesDGX agent

arXiv:2606.11120v1 Announce Type: new Abstract: We recast pass evaluation in football (soccer) as a Monte Carlo Tree Search (MCTS)-like evaluation problem whose components mostly exist in the literatu

One Step Closer to Ground Truth: A Multi-Scale Residual-Aware Representation Learning Pipeline for Predicting Time Series Data

Model ReleasesDGX agent

arXiv:2606.10678v1 Announce Type: new Abstract: Transformer-based models have emerged as leading paradigms in time-series forecasting in recent years, employing self-attention mechanisms to capture lo

PreAct-Bench: Benchmarking Predictive Monitoring in LLMs

Model ReleasesDGX agent

arXiv:2606.09890v1 Announce Type: cross Abstract: Large language models (LLMs) are increasingly deployed as autonomous agents capable of executing multi-step action trajectories toward a given objecti

Quo Vadis, Visual In-Context Learning? A Unified Benchmark Across Domains and Tasks

Model ReleasesDGX agent

arXiv:2606.10967v1 Announce Type: new Abstract: Visual in-context learning has been proposed as a pathway towards dynamic models that can generate predictions based on a provided context and thereby c

RoboGPT-R1: Enhancing Robot Task Planning with Reinforcement Learning

Model ReleasesDGX agent

arXiv:2510.14828v3 Announce Type: replace Abstract: Improving the reasoning capabilities of embodied agents is crucial for robots to complete complex human instructions in long-view manipulation tasks

SAFE: An LLM-as-Verifier Framework for Evidence-Grounded Multi-Hop Reasoning

Model ReleasesDGX agent

arXiv:2604.01993v2 Announce Type: replace-cross Abstract: Multi-hop QA benchmarks often reward Large Language Models (LLMs) for spurious correctness, where models reach correct answers through invalid

Segmentation-Driven Monocular Shape from Polarization based on Physical Model

ApplicationsDGX agent

arXiv:2601.04776v2 Announce Type: replace Abstract: Monocular shape-from-polarization (SfP) leverages the intrinsic relationship between light polarization properties and surface geometry to recover s

Superficial Beliefs in LLM Decision-Making

ResearchDGX agent

arXiv:2606.11016v1 Announce Type: new Abstract: We ask whether large language models (LLMs) merely imitate rationales when choosing between two options, or whether their choices reflect a systematic u

U-TTT: Towards Generalizable PET Image Denoising via Test-Time Training

Model ReleasesDGX agent

arXiv:2606.11032v1 Announce Type: new Abstract: Existing deep learning models for Positron Emission Tomography (PET) image denoising often suffer from severe performance degradation under distribution

Validation-Stage Combinatorial Fusion Analysis for Imbalanced Credit-Card Fraud Detection

Model ReleasesDGX agent

arXiv:2606.10393v1 Announce Type: new Abstract: Credit-card fraud detection is difficult because fraudulent transactions are rare, costly, and unevenly distributed. Strong gradient-boosted tree models

WebChallenger: A Reliable and Efficient Generalist Web Agent

Model ReleasesDGX agent

arXiv:2606.10423v1 Announce Type: new Abstract: Autonomous web navigation remains challenging for LLM agents, and the strongest generalist systems rely on proprietary reasoning models whose inference

What Fits (Into Few Tokens) Doesn't Overfit: Compression and Generalization in ML Research Agents

Model ReleasesDGX agent

arXiv:2606.11045v1 Announce Type: new Abstract: Reusing a held-out benchmark adaptively should, in principle, invite overfitting. Yet benchmark-driven machine learning (ML) has produced surprisingly l

← Previous
1…277278279280281…1034
Next →