AI Wiki
TimelineEvolutionGraphStatusAsk wiki
Live from Git
AI Wiki
TimelineEvolutionGraphStatusAsk wiki
Live from Git
Filter entries
Categories
  • All entries86,428
  • Agents7,398
  • Applications5,301
  • Concepts5
  • Hardware1,785
  • Industry6,113
  • Local Ai4,833
  • Model Releases23,177
  • Research19,713
  • Safety13,092
  • Syntheses17
  • Tools1,670
  • Tutorials3,324

Source
HumanDGX agent

Content type
AllBlogX PostPaperYouTubeRedditGitHub
Clear filters
Categories
  • All entries86,428
  • Agents7,398
  • Applications5,301
  • Concepts5
  • Hardware1,785
  • Industry6,113
  • Local Ai4,833
  • Model Releases23,177
  • Research19,713
  • Safety13,092
  • Syntheses17
  • Tools1,670
  • Tutorials3,324

Source
HumanDGX agent

86,428Total entries
1Added by human
86,427Found by agent
12Categories

Knowledge catalogue

Search: “models”

GridTimelineEvolution
62,016 results
18 May 2026

Greedy or not, here I come: Language production under vocabulary constraints in humans and resource-rational models

ApplicationsDGX agent

arXiv:2605.15365v1 Announce Type: new Abstract: Communicating using only a limited vocabulary is a common but challenging cognitive phenomenon, requiring an ideal communicator to plan carefully to opt

GRLO: Towards Generalizable Reinforcement Learning in Open-Ended Environments from Zero

Model ReleasesDGX agent

arXiv:2605.15464v1 Announce Type: cross Abstract: Post-training has become a crucial step for unlocking the capabilities of large language models, with reinforcement learning (RL) emerging as a critic

Looped SSMs: Depth-Recurrence and Input Reshaping for Time Series Classification

Model ReleasesDGX agent
Content type
AllBlogX PostPaperYouTubeRedditGitHub
Clear filters

arXiv:2605.16048v1 Announce Type: cross Abstract: State Space Models (SSMs) are inherently recurrent along the sequence dimension, yet depth-recurrence - reusing the same block repeatedly across layer

MAgSeg: Segmentation of Agricultural Landscapes in High-Resolution Satellite Imagery using Multimodal Large Language Models

SafetyDGX agent

arXiv:2605.16179v1 Announce Type: new Abstract: Agricultural landscape segmentation in the Global South is challenging as it is characterized by fragmented plots, high intra-class variance, and a scar

NIMO Controller: a self-driving laboratory orchestrator based on the Model Context Protocol

AgentsDGX agent

arXiv:2605.15227v1 Announce Type: new Abstract: Self-driving laboratories (SDLs) have attracted increasing attention as a means of accelerating scientific discovery; however, developing SDL software r

Preprocessing Algorithm Leveraging Geometric Modeling for Scale Correction in Hyperspectral Images for Improved Unmixing Performance

ApplicationsDGX agent

arXiv:2508.08431v3 Announce Type: replace-cross Abstract: Spectral variability significantly impacts the accuracy and convergence of hyperspectral unmixing algorithms. Many methods address complex spe

Second-Order Multi-Level Variance Correction for Modality Competition in Multimodal Models

SafetyDGX agent

arXiv:2605.16165v1 Announce Type: cross Abstract: Autoregressive next-token training offers a unified formulation for image generation and text understanding, but it also creates strong modality compe

STS: Efficient Sparse Attention with Speculative Token Sparsity

Model ReleasesDGX agent

arXiv:2605.15508v1 Announce Type: cross Abstract: The quadratic complexity of attention imposes severe memory and computational bottlenecks on Large Language Model (LLM) inference. This challenge is p

Towards Trustworthy and Explainable AI for Perception Models: From Concept to Prototype Vehicle Deployment

SafetyDGX agent

arXiv:2605.16087v1 Announce Type: cross Abstract: Deep Neural Networks have become the dominant solution for Autonomous Driving perception, but their opacity conflicts with emerging Trustworthy AI gui

15 May 2026

Claude Code's product lead talks usage limits, transparency, and the 'lean harness'

Model ReleasesDGX agent

Claude Code's product lead addresses how Anthropic tunes the 'harness' (the structural layer around the model) for each new model release to optimize performance and reduce verbosity. The company comm

ClickRemoval: An Interactive Open-Source Tool for Object Removal in Diffusion Models

ResearchDGX agent

arXiv:2605.14461v1 Announce Type: new Abstract: Existing object removal tools often rely on manual masks or text prompts, making precise removal difficult for non-expert users in complex scenes and of

Correctness-Aware Repository Filtering Under Maximum Effective Context Window Constraints

Model ReleasesDGX agent

arXiv:2605.14362v1 Announce Type: cross Abstract: Context window efficiency is a practical constraint in large language model (LLM)-based developer tools. Paulsen [12] shows that all tested models deg

From Street View to Visual Network: Mapping the Visibility of Urban Landmarks with Vision-Language Models

ApplicationsDGX agent

arXiv:2505.11809v3 Announce Type: replace Abstract: Visibility analysis in urban planning has traditionally relied on line-of-sight (LoS) simulations, which capture geometric occlusion. However, these

PacTure: Efficient PBR Texture Generation on Packed Views with Visual Autoregressive Models

ResearchDGX agent

arXiv:2505.22394v2 Announce Type: replace Abstract: We present PacTure, a novel framework for generating physically-based rendering (PBR) material textures for an untextured 3D mesh from a text descri

TiTok: Transfer Token-level Knowledge via Contrastive Excess to Transplant LoRA

Model ReleasesDGX agent

arXiv:2510.04682v3 Announce Type: replace-cross Abstract: Large Language Models (LLMs) are widely applied in real world scenarios, yet fine-tuning them comes with significant computational and storage

14 May 2026

A Hybrid Tucker-LSTM Tensor Network Model for SOC Prediction in Electric Vehicles

ApplicationsDGX agent

arXiv:2605.13200v1 Announce Type: new Abstract: Accurate state of charge estimation is critical for the success of electric vehicle battery management strategies, but it is well known that conventiona

AI Safety Landscape for Large Language Models: Taxonomy, State-of-the-art, and Future Directions

SafetyDGX agent

arXiv:2408.12935v4 Announce Type: replace Abstract: AI Safety is an emerging area of critical importance to the safe adoption and deployment of AI systems. With the rapid proliferation of AI and espec

An Agentic AI Framework with Large Language Models and Chain-of-Thought for UAV-Assisted Logistics Scheduling with Mobile Edge Computing

Local AiDGX agent

arXiv:2605.13221v1 Announce Type: new Abstract: In cloud manufacturing, unmanned aerial vehicles (UAVs) can support both product collection and mobile edge computing (MEC). This joint operation forms

Causality-Aware End-to-End Autonomous Driving via Ego-Centric Joint Scene Modeling

SafetyDGX agent

arXiv:2605.13646v1 Announce Type: cross Abstract: End-to-end autonomous driving, which bypasses traditional modular pipelines by directly predicting future trajectories from sensor inputs, has recentl

ChipMATE: Multi-Agent Training via Reinforcement Learning for Enhanced RTL Generation

Model ReleasesDGX agent

arXiv:2605.12857v1 Announce Type: cross Abstract: Existing API-based agentic systems for RTL code generation are fundamentally misaligned with industrial practice: they assume a golden testbench is av

CodeClash: Benchmarking Goal-Oriented Software Engineering

Model ReleasesDGX agent

arXiv:2511.00839v2 Announce Type: replace-cross Abstract: Current benchmarks for coding evaluate language models (LMs) on concrete, well-specified tasks such as fixing specific bugs or writing targete

GenCape: Structure-Inductive Generative Modeling for Category-Agnostic Pose Estimation

Local AiDGX agent

arXiv:2605.13151v1 Announce Type: new Abstract: Category-agnostic pose estimation (CAPE) aims to localize keypoints on query images from arbitrary categories, using only a few annotated support exampl

Learning a Continue-Thinking Token for Enhanced Test-Time Scaling

Model ReleasesDGX agent

arXiv:2506.11274v2 Announce Type: replace-cross Abstract: Test-time scaling has emerged as an effective approach for improving language model performance by utilizing additional compute at inference t

Learning to Optimize Radiotherapy Plans via Fluence Maps Diffusion Model Generation and LSTM-based Optimization

ResearchDGX agent

arXiv:2605.13713v1 Announce Type: new Abstract: Volumetric Modulated Arc Therapy (VMAT) is a cornerstone of modern radiation therapy, enabling highly conformal tumor irradiation and healthy-tissue spa

LLMs as annotators of credibility assessment in Danish asylum decisions: evaluating classification performance and errors beyond aggregated metrics

Model ReleasesDGX agent

arXiv:2605.13412v1 Announce Type: cross Abstract: Off-the-shelf large language models (LLMs) are increasingly used to automate text annotation, yet their effectiveness remains underexplored for underr

Locale-Conditioned Few-Shot Prompting Mitigates Demonstration Regurgitation in On-Device PII Substitution with Small Language Models

Local AiDGX agent

arXiv:2605.13538v1 Announce Type: cross Abstract: Personally Identifiable Information (PII) redaction usually replaces detected entities with placeholder tokens such as [PERSON], destroying the downst

Make-It-Poseable: Feed-forward Latent Posing Model for 3D Characters

ResearchDGX agent

arXiv:2512.16767v2 Announce Type: replace Abstract: Posing 3D characters is a fundamental task in computer graphics. However, existing paradigms, ranging from traditional auto-rigging to recent pose-c

Mix, Don't Tune: Bilingual Pre-Training Outperforms Hyperparameter Search in Data-Constrained Settings

ResearchDGX agent

arXiv:2605.13225v1 Announce Type: new Abstract: For most languages of the world, language model pre-training operates in a data-constrained regime where models must repeat their training data many tim

Multimodal Hidden Markov Models for Persistent Emotional State Tracking

ResearchDGX agent

arXiv:2605.12838v1 Announce Type: new Abstract: Tracking an interpretable emotional arc of a conversation via the sentiment of individual utterances processed as a whole is central to both understandi

Repurposing Image Diffusion Models for Training-Free Music Style Transfer on Mel-spectrograms

ResearchDGX agent

arXiv:2411.15913v4 Announce Type: replace-cross Abstract: Music style transfer blends source structure with reference style to enable personalized music creation. However, existing zero-shot methods o

Scale-Gest: Scalable Model-Space Synthesis and Runtime Selection for On-Device Gesture Detection

Local AiDGX agent

arXiv:2605.12506v1 Announce Type: cross Abstract: Realizing on-device ML-based gesture detection under tight real-time performance, energy and memory constraints is challenging, especially when consid

Senses Wide Shut: A Representation-Action Gap in Omnimodal LLMs

Model ReleasesDGX agent

arXiv:2605.13737v1 Announce Type: new Abstract: When an omnimodal large language model accepts a question whose textual premise contradicts what it actually sees or hears, does the failure lie in perc

Table-R1: Region-based Reinforcement Learning for Table Understanding

Model ReleasesDGX agent

arXiv:2505.12415v3 Announce Type: replace-cross Abstract: Tables present unique challenges for language models due to their structured row-column interactions, necessitating specialized approaches for

The Mechanism of Weak-to-Strong Generalization: Feature Elicitation from Latent Knowledge

ResearchDGX agent

arXiv:2605.12908v1 Announce Type: cross Abstract: Weak-to-strong (W2S) generalization, in which a strong model is fine-tuned on outputs of a weaker, task-specialized model, has been proposed as an app

Training LLMs with Reinforcement Learning for Intent-Aware Personalized Question Answering

Model ReleasesDGX agent

arXiv:2605.12645v1 Announce Type: cross Abstract: Effective personalized question answering (PQA) in language models requires grounding responses in the user's underlying intent, where intent refers t

13 May 2026

Agent-BRACE: Decoupling Beliefs from Actions in Long-Horizon Tasks via Verbalized State Uncertainty

Model ReleasesDGX agent

arXiv:2605.11436v1 Announce Type: new Abstract: Large language models (LLMs) are increasingly deployed on long-horizon tasks in partially observable environments, where they must act while inferring a

Beyond Prediction: Interval Neural Networks for Uncertainty-Aware System Identification

Model ReleasesDGX agent

arXiv:2605.11460v1 Announce Type: new Abstract: System identification (SysID) is critical for modeling dynamical systems from experimental data, yet traditional approaches often fail to capture nonlin

BitLM: Unlocking Multi-Token Language Generation with Bitwise Continuous Diffusion

ResearchDGX agent

arXiv:2605.11577v1 Announce Type: new Abstract: Autoregressive language models generate text one token at a time, yet natural language is inherently structured in multi-token units, including phrases,

Built a Chrome extension that talks directly with your local Ollama models

Local AiDGX agent

A Chrome extension that communicates directly with local Ollama instances without sending data to external servers , enabling users to submit messages through a popup interface with responses streamin

DemaFormer: Damped Exponential Moving Average Transformer with Energy-Based Modeling for Temporal Language Grounding

Local AiDGX agent

arXiv:2312.02549v2 Announce Type: replace-cross Abstract: Temporal Language Grounding seeks to localize video moments that semantically correspond to a natural language query. Recent advances employ t

Efficient and Adaptive Human Activity Recognition via LLM Backbones

Model ReleasesDGX agent

arXiv:2605.12019v1 Announce Type: new Abstract: Human Activity Recognition (HAR) is a core task in pervasive computing systems, where models must operate under strict computational constraints while r

EvoNav: Evolutionary Reward Function Design for Robot Navigation with Large Language Models

SafetyDGX agent

arXiv:2605.11859v1 Announce Type: new Abstract: Robot navigation is a crucial task with applications to social robots in dynamic human environments. While Reinforcement Learning (RL) has shown great p

FLARE: Adaptive Multi-Dimensional Reputation for Robust Client Reliability in Federated Learning

Model ReleasesDGX agent

arXiv:2511.14715v3 Announce Type: replace Abstract: Federated learning (FL) enables collaborative model training while preserving data privacy. However, it remains vulnerable to malicious clients who

Investigating simple target-covariate relationships for Chronos-2 and TabPFN-TS

Model ReleasesDGX agent

arXiv:2605.12200v1 Announce Type: new Abstract: Time Series Foundation Models (TSFMs) have recently achieved state-of-the-art performance, often outperforming supervised models in zero-shot settings.

Modeling Narrative Structure in Latin Epic Poetry with Automatically Generated Story Grammars

ResearchDGX agent

arXiv:2502.12276v2 Announce Type: replace Abstract: Computational methods for analyzing prose and poetry utilize word embeddings and other abstract representations that sometimes obscure context-rich

Multimodal Abstractive Summarization of Instructional Videos with Vision-Language Models

SafetyDGX agent

arXiv:2605.11959v1 Announce Type: cross Abstract: Multimodal video summarization requires visual features that align semantically with language generation. Traditional approaches rely on CNN features

Quantifying the Reconstructability of Astrophysical Methods with Large Language Models and Information Theory: A Case Study in Spectral Reconstruction

ApplicationsDGX agent

arXiv:2605.11154v1 Announce Type: cross Abstract: Modern astrophysical studies rely heavily on complex data analysis pipelines; however, published descriptions often lack the detail required for compu

See the past: Time-Reversed Scene Reconstruction from Thermal Traces Using Visual Language Models

ResearchDGX agent

arXiv:2510.05408v2 Announce Type: replace Abstract: Recovering the past from present observations is an intriguing challenge with potential applications in forensics and scene analysis. Thermal imagin

Selection, Not Fusion: Radar-Modulated State Space Models for Radar-Camera Depth Estimation

ApplicationsDGX agent

arXiv:2605.11840v1 Announce Type: new Abstract: Radar-camera depth estimation must turn an ultra-sparse, all-weather, metric radar signal into a dense per-pixel depth map. Existing methods -- concaten

Self-Consistent Latent Reasoning: Long Latent Sequence Reasoning for Vision-Language Model

ApplicationsDGX agent

arXiv:2605.12163v1 Announce Type: new Abstract: In language reasoning, longer chains of thought consistently yield better performance, which naturally suggests that visual latent reasoning may likewis

Smoothness Errors in Dynamics Models and How to Avoid Them

TutorialsDGX agent

arXiv:2602.05352v2 Announce Type: replace Abstract: Modern neural networks have shown promise for solving partial differential equations over surfaces, often by discretizing the surface as a mesh and

TabDLM: Free-Form Tabular Data Generation via Joint Numerical-Language Diffusion

ApplicationsDGX agent

arXiv:2602.22586v2 Announce Type: replace-cross Abstract: Synthetic tabular data generation has attracted growing attention due to its importance for data augmentation, foundation models, and privacy.

Transferable Delay-Aware Reinforcement Learning via Implicit Causal Graph Modeling

SafetyDGX agent

arXiv:2605.12312v1 Announce Type: new Abstract: Random delays weaken the temporal correspondence between actions and subsequent state feedback, making it difficult for agents to identify the true prop

Vision2Code: A Multi-Domain Benchmark for Evaluating Image-to-Code Generation

Model ReleasesDGX agent

arXiv:2605.11307v1 Announce Type: new Abstract: Image-to-code generation tests whether a vision-language model (VLM) can recover the structure of an image enough to express it as executable code. Exis

ZeroIDIR: Zero-Reference Illumination Degradation Image Restoration with Perturbed Consistency Diffusion Models

SafetyDGX agent

arXiv:2605.11435v1 Announce Type: new Abstract: In this paper, we propose a zero-reference diffusion-based framework, named ZeroIDIR, for illumination degradation image restoration, which decouples th

12 May 2026

AdaptSplat: Adapting Vision Foundation Models for Feed-Forward 3D Gaussian Splatting

ResearchDGX agent

arXiv:2605.10239v1 Announce Type: new Abstract: This work explores a simple yet powerful lightweight adapter design for feed-forward 3D Gaussian Splatting (3DGS). Existing methods typically apply comp

An Integrative Genome-Scale Metabolic Modeling and Machine Learning Framework for Predicting and Optimizing Single-Cell Protein Production in Saccharomyces cerevisiae

ApplicationsDGX agent

arXiv:2603.25561v2 Announce Type: replace Abstract: Saccharomyces cerevisiae is increasingly recognised as a key source for single-cell protein (SCP) production, a rising solution to global protein-su

AR-VLA: True Autoregressive Action Expert for Vision-Language-Action Models

ResearchDGX agent

arXiv:2603.10126v2 Announce Type: replace-cross Abstract: We propose a standalone autoregressive (AR) Action Expert that generates actions as a continuous causal sequence while conditioning on refresh

Can LLMs Predict Polymer Physics Just by Reading Synthesis and Processing Prose?

Model ReleasesDGX agent

arXiv:2605.08255v1 Announce Type: cross Abstract: Can large language models predict physical and mechanical polymer properties simply by reading unstructured scientific prose? Polymer performance is r

Decentralized Conformal Novelty Detection via Quantized Model Exchange

ResearchDGX agent

arXiv:2605.08263v1 Announce Type: cross Abstract: This work studies decentralized novelty detection with global false discovery rate (FDR) control across heterogeneous composite null distributions, wi

← Previous
1…261262263264265…1034
Next →