AI Wiki
TimelineEvolutionGraphStatusAsk wiki
Live from Git
AI Wiki
TimelineEvolutionGraphStatusAsk wiki
Live from Git
Filter entries
Categories
  • All entries86,993
  • Agents7,449
  • Applications5,325
  • Concepts5
  • Hardware1,798
  • Industry6,136
  • Local Ai4,859
  • Model Releases23,375
  • Research19,835
  • Safety13,176
  • Syntheses17
  • Tools1,670
  • Tutorials3,348

Source
HumanDGX agent

Content type
AllBlogX PostPaperYouTubeRedditGitHub
Clear filters
Categories
  • All entries86,993
  • Agents7,449
  • Applications5,325
  • Concepts5
  • Hardware1,798
  • Industry6,136
  • Local Ai4,859
  • Model Releases23,375
  • Research19,835
  • Safety13,176
  • Syntheses17
  • Tools1,670
  • Tutorials3,348

Source
HumanDGX agent

86,993Total entries
1Added by human
86,992Found by agent
12Categories

Knowledge catalogue

Search: “models”

GridTimelineEvolution
62,477 results
15 Apr 2026

Think Through Uncertainty: Improving Long-Form Generation Factuality via Reasoning Calibration

ResearchDGX agent

arXiv:2604.12046v1 Announce Type: new Abstract: Large language models (LLMs) often hallucinate in long-form generation. Existing approaches mainly improve factuality through post-hoc revision or reinf

ToxiTrace: Gradient-Aligned Training for Explainable Chinese Toxicity Detection

Model ReleasesDGX agent

arXiv:2604.12321v1 Announce Type: new Abstract: Existing Chinese toxic content detection methods mainly target sentence-level classification but often fail to provide readable and contiguous toxic evi

Transforming External Knowledge into Triplets for Enhanced Retrieval in RAG of LLMs

Model ReleasesDGX agent
Content type
AllBlogX PostPaperYouTubeRedditGitHub
Clear filters

arXiv:2604.12610v1 Announce Type: new Abstract: Retrieval-Augmented Generation (RAG) mitigates hallucination in large language models (LLMs) by incorporating external knowledge during generation. Howe

ZIB ZIT hand over?

Local AiDGX agent

This r/StableDiffusion Reddit thread likely discusses the topic of hand generation quality when using ZIB and ZIT — two AI image generation models from the Z-Image ecosystem used in Stable Diffusion a

14 Apr 2026

A Complete Decomposition of KL Error using Refined Information and Mode Interaction Selection

ApplicationsDGX agent

arXiv:2410.11964v2 Announce Type: replace Abstract: The log-linear model has received a significant amount of theoretical attention in previous decades and remains the fundamental tool used for learni

A Hybrid Intelligent Framework for Uncertainty-Aware Condition Monitoring of Industrial Systems

Model ReleasesDGX agent

arXiv:2604.09932v1 Announce Type: cross Abstract: Hybrid approaches that combine data-driven learning with physics-based insight have shown promise for improving the reliability of industrial conditio

A Unified Theory of Sparse Dictionary Learning in Mechanistic Interpretability: Piecewise Biconvexity and Spurious Minima

TutorialsDGX agent

arXiv:2512.05534v4 Announce Type: replace-cross Abstract: As AI models achieve remarkable capabilities across diverse domains, understanding what representations they learn and how they encode concept

A Weak Penalty Neural ODE for Learning Chaotic Dynamics from Noisy Time Series

Model ReleasesDGX agent

arXiv:2511.06609v3 Announce Type: replace Abstract: The accurate forecasting of complex, high-dimensional dynamical systems from observational data is a fundamental task across numerous scientific and

ADD for Multi-Bit Image Watermarking

Model ReleasesDGX agent

arXiv:2604.11491v1 Announce Type: cross Abstract: As generative models enable rapid creation of high-fidelity images, societal concerns about misinformation and authenticity have intensified. A promis

Agentic Aggregation for Parallel Scaling of Long-Horizon Agentic Tasks

Model ReleasesDGX agent

arXiv:2604.11753v1 Announce Type: new Abstract: We study parallel test-time scaling for long-horizon agentic tasks such as agentic search and deep research, where multiple rollouts are generated in pa

Ambivalence/Hesitancy Recognition in Videos for Personalized Digital Health Interventions

ResearchDGX agent

arXiv:2604.11730v1 Announce Type: new Abstract: Using behavioural science, health interventions focus on behaviour change by providing a framework to help patients acquire and maintain healthy habits

C-ReD: A Comprehensive Chinese Benchmark for AI-Generated Text Detection Derived from Real-World Prompts

Model ReleasesDGX agent

arXiv:2604.11796v1 Announce Type: cross Abstract: Recently, large language models (LLMs) are capable of generating highly fluent textual content. While they offer significant convenience to humans, th

Can Multi-Modal LLMs Provide Live Step-by-Step Task Guidance?

Model ReleasesDGX agent

arXiv:2511.21998v2 Announce Type: replace Abstract: Multi-modal Large Language Models (LLM) have advanced conversational abilities but struggle with providing live, interactive step-by-step guidance,

CARO: Chain-of-Analogy Reasoning Optimization for Robust Content Moderation

Model ReleasesDGX agent

arXiv:2604.10504v1 Announce Type: new Abstract: Current large language models (LLMs), even those explicitly trained for reasoning, often struggle with ambiguous content moderation cases due to mislead

CoPS: Conditional Prompt Synthesis for Zero-Shot Anomaly Detection

SafetyDGX agent

arXiv:2508.03447v2 Announce Type: replace Abstract: Recently, large pre-trained vision-language models have shown remarkable performance in zero-shot anomaly detection (ZSAD). With fine-tuning on a si

Counting to Four is still a Chore for VLMs

Model ReleasesDGX agent

arXiv:2604.10039v1 Announce Type: new Abstract: Vision--language models (VLMs) have achieved impressive performance on complex multimodal reasoning tasks, yet they still fail on simple grounding skill

Dead Cognitions: A Census of Misattributed Insights

Model ReleasesDGX agent

arXiv:2604.10288v1 Announce Type: new Abstract: This essay identifies a failure mode of AI chat systems that we term attribution laundering: the model performs substantive cognitive work and then rhet

Decompose, Mix, Adapt: A Unified Framework for Parameter-Efficient Neural Network Recombination and Compression

Model ReleasesDGX agent

arXiv:2603.27383v2 Announce Type: replace Abstract: Parameter Recombination (PR) methods aim to efficiently compose the weights of a neural network for applications like Parameter-Efficient FineTuning

Discrete Flow Maps

SafetyDGX agent

arXiv:2604.09784v1 Announce Type: cross Abstract: The sequential nature of autoregressive next-token prediction imposes a fundamental speed limit on large language models. While continuous flow models

DiSPA: Differential Substructure-Pathway Attention for Drug Response Prediction

Model ReleasesDGX agent

arXiv:2601.14346v2 Announce Type: replace-cross Abstract: Accurate prediction of drug response in precision medicine requires models that capture how specific chemical substructures interact with cell

DistDF: Time-Series Forecasting Needs Joint-Distribution Wasserstein Alignment

SafetyDGX agent

arXiv:2510.24574v2 Announce Type: replace-cross Abstract: Training time-series forecasting models requires aligning the conditional distribution of model forecasts with that of the label sequence. The

ERNIE Image released

Model ReleasesDGX agent

ERNIE Image is an open-source text-to-image generation model developed by Baidu, built on a single-stream Diffusion Transformer (DiT) paired with a lightweight Prompt Enhancer that expands brief user

Escaping the Context Bottleneck: Active Context Curation for LLM Agents via Reinforcement Learning

Model ReleasesDGX agent

arXiv:2604.11462v1 Announce Type: new Abstract: Large Language Models (LLMs) struggle with long-horizon tasks due to the 'context bottleneck' and the 'lost-in-the-middle' phenomenon, where accumulated

FedKLPR: KL-Guided Pruning-Aware Federated Learning for Person Re-Identification

Model ReleasesDGX agent

arXiv:2508.17431v3 Announce Type: replace-cross Abstract: Person re-identification (re-ID) is a fundamental task in intelligent surveillance and public safety. Federated learning (FL) provides a priva

FedQUIT: On-Device Federated Unlearning via a Quasi-Competent Virtual Teacher

Local AiDGX agent

arXiv:2408.07587v4 Announce Type: replace Abstract: Federated Learning (FL) enables the collaborative training of machine learning models without requiring centralized collection of user data. To comp

FinTrace: Holistic Trajectory-Level Evaluation of LLM Tool Calling for Long-Horizon Financial Tasks

Model ReleasesDGX agent

arXiv:2604.10015v1 Announce Type: new Abstract: Recent studies demonstrate that tool-calling capability enables large language models (LLMs) to interact with external environments for long-horizon fin

FlexMS is a flexible framework for benchmarking deep learning-based mass spectrum prediction tools in metabolomics

Model ReleasesDGX agent

arXiv:2602.22822v2 Announce Type: replace Abstract: The identification and property prediction of chemical molecules is of central importance in the advancement of drug discovery and material science,

From Scalars to Tensors: Declared Losses Recover Epistemic Distinctions That Neutrosophic Scalars Cannot Express

Model ReleasesDGX agent

arXiv:2604.09602v1 Announce Type: new Abstract: Leyva-Vazquez and Smarandache (2025) demonstrated that neutrosophic T/I/F evaluation, where Truth, Indeterminacy, and Falsity are independent dimensions

GeoFormer: A Lightweight Swin Transformer for Joint Building Height and Footprint Estimation from Sentinel Imagery

Model ReleasesDGX agent

arXiv:2602.09932v2 Announce Type: replace Abstract: Building height (BH) and footprint (BF) are fundamental urban morphological parameters required by climate modelling, disaster-risk assessment, and

GoT-R1: Unleashing Reasoning Capability of MLLM for Visual Generation with Reinforcement Learning

Model ReleasesDGX agent

arXiv:2505.17022v2 Announce Type: replace-cross Abstract: Visual generation models have made remarkable progress in creating realistic images from text prompts, yet struggle with complex prompts that

Hardware Utilization and Inference Performance of Edge Object Detection Under Fault Injection

HardwareDGX agent

arXiv:2604.09631v1 Announce Type: cross Abstract: As deep learning models are deployed on resource constrained edge platforms in autonomous driving systems, reli able knowledge of hardware behavior un

HDR Video Generation via Latent Alignment with Logarithmic Encoding

SafetyDGX agent

arXiv:2604.11788v1 Announce Type: new Abstract: High dynamic range (HDR) imagery offers a rich and faithful representation of scene radiance, but remains challenging for generative models due to its m

Improving LLM Unlearning Robustness via Random Perturbations

ResearchDGX agent

arXiv:2501.19202v5 Announce Type: replace Abstract: Here, we show that current LLM unlearning methods inherently reduce models' robustness, causing them to misbehave even when a single non-adversarial

LARY: A Latent Action Representation Yielding Benchmark for Generalizable Vision-to-Action Alignment

Model ReleasesDGX agent

arXiv:2604.11689v1 Announce Type: new Abstract: While the shortage of explicit action data limits Vision-Language-Action (VLA) models, human action videos offer a scalable yet unlabeled data source. A

Learning Long-term Motion Embeddings for Efficient Kinematics Generation

TutorialsDGX agent

arXiv:2604.11737v1 Announce Type: new Abstract: Understanding and predicting motion is a fundamental component of visual intelligence. Although modern video models exhibit strong comprehension of scen

LIDARLearn: A Unified Deep Learning Library for 3D Point Cloud Classification, Segmentation, and Self-Supervised Representation Learning

Model ReleasesDGX agent

arXiv:2604.10780v1 Announce Type: new Abstract: Three-dimensional (3D) point cloud analysis has become central to applications ranging from autonomous driving and robotics to forestry and ecological m

LiveCLKTBench: Towards Reliable Evaluation of Cross-Lingual Knowledge Transfer in Multilingual LLMs

Model ReleasesDGX agent

arXiv:2511.14774v3 Announce Type: replace-cross Abstract: Evaluating cross-lingual knowledge transfer in large language models is challenging, as correct answers in a target language may arise either

LLMs for Text-Based Exploration and Navigation Under Partial Observability

Model ReleasesDGX agent

arXiv:2604.09604v1 Announce Type: new Abstract: Exploration and goal-directed navigation in unknown layouts are central to inspection, logistics, and search-and-rescue. We ask whether large language m

Lung Cancer Detection Using Deep Learning

ResearchDGX agent

arXiv:2604.10765v1 Announce Type: cross Abstract: Lung cancer, the second leading cause of cancer-related deaths, is primarily linked to long-term tobacco smoking (85% of cases). Surprisingly, 10-15%

Masked Contrastive Pre-Training Improves Music Audio Key Detection

ResearchDGX agent

arXiv:2604.10021v1 Announce Type: cross Abstract: Self-supervised music foundation models underperform on key detection, which requires pitch-sensitive representations. In this work, we present the fi

MCAT: Scaling Many-to-Many Speech-to-Text Translation with MLLMs to 70 Languages

Model ReleasesDGX agent

arXiv:2512.01512v2 Announce Type: replace Abstract: Multimodal Large Language Models (MLLMs) have achieved great success in Speech-to-Text Translation (S2TT) tasks. However, current research is constr

Measuring What Matters!! Assessing Therapeutic Principles in Mental-Health Conversation

Model ReleasesDGX agent

arXiv:2604.05795v2 Announce Type: replace Abstract: The increasing use of large language models in mental health applications calls for principled evaluation frameworks that assess alignment with psyc

MPAC: A Multi-Principal Agent Coordination Protocol for Interoperable Multi-Agent Collaboration

Model ReleasesDGX agent

arXiv:2604.09744v1 Announce Type: cross Abstract: The AI agent ecosystem has converged on two protocols: the Model Context Protocol (MCP) for tool invocation and Agent-to-Agent (A2A) for single-princi

Network Effects and Agreement Drift in LLM Debates

ResearchDGX agent

arXiv:2604.11312v1 Announce Type: cross Abstract: Large Language Models (LLMs) have demonstrated an unprecedented ability to simulate human-like social behaviors, making them useful tools for simulati

One of my passions is that education should be dispersed freely and as widely as possible, especially for technologies as dynamic and crucia…

Model ReleasesDGX agent

One of my passions is that education should be dispersed freely and as widely as possible, especially for technologies as dynamic and crucial as LLMs/AI. I'm proud to have friends who would disown me

Panoptic Pairwise Distortion Graph

Model ReleasesDGX agent

arXiv:2604.11004v1 Announce Type: cross Abstract: In this work, we introduce a new perspective on comparative image assessment by representing an image pair as a structured composition of its regions.

Point2Pose: Occlusion-Recovering 6D Pose Tracking and 3D Reconstruction for Multiple Unknown Objects Via 2D Point Trackers

Model ReleasesDGX agent

arXiv:2604.10415v1 Announce Type: new Abstract: We present Point2Pose, a model-free method for causal 6D pose tracking of multiple rigid objects from monocular RGB-D video. Initialized only from spars

PoTable: Towards Systematic Thinking via Plan-then-Execute Stage Reasoning on Tables

AgentsDGX agent

arXiv:2412.04272v5 Announce Type: replace-cross Abstract: In recent years, table reasoning has garnered substantial research interest, particularly regarding its integration with Large Language Models

Preference Learning Unlocks LLMs' Psycho-Counseling Skills

ResearchDGX agent

arXiv:2502.19731v2 Announce Type: replace Abstract: Applying large language models (LLMs) to assist in psycho-counseling is an emerging and meaningful approach, driven by the significant gap between p

SignReasoner: Compositional Reasoning for Complex Traffic Sign Understanding via Functional Structure Units

Model ReleasesDGX agent

arXiv:2604.10436v1 Announce Type: new Abstract: Accurate semantic understanding of complex traffic signs-including those with intricate layouts, multi-lingual text, and composite symbols-is critical f

Specificity-aware reinforcement learning for fine-grained open-world classification

TutorialsDGX agent

arXiv:2603.03197v3 Announce Type: replace Abstract: Classifying fine-grained visual concepts under open-world settings, i.e., without a predefined label set, demands models to be both accurate and spe

StarVLA-alpha: Reducing Complexity in Vision-Language-Action Systems

Model ReleasesDGX agent

arXiv:2604.11757v1 Announce Type: cross Abstract: Vision-Language-Action (VLA) models have recently emerged as a promising paradigm for building general-purpose robotic agents. However, the VLA landsc

Structured State-Space Regularization for Compact and Generation-Friendly Image Tokenization

TutorialsDGX agent

arXiv:2604.11089v1 Announce Type: new Abstract: Image tokenizers are central to modern vision models as they often operate in latent spaces. An ideal latent space must be simultaneously compact and ge

Thinking Fast, Thinking Wrong: Intuitiveness Modulates LLM Counterfactual Reasoning in Policy Evaluation

Model ReleasesDGX agent

arXiv:2604.10511v1 Announce Type: new Abstract: Large language models (LLMs) are increasingly used for causal and counterfactual reasoning, yet their reliability in real-world policy evaluation remain

Tracking High-order Evolutions via Cascading Low-rank Fitting

Model ReleasesDGX agent

arXiv:2604.10980v1 Announce Type: new Abstract: Diffusion models have become the de facto standard for modern visual generation, including well-established frameworks such as latent diffusion and flow

Triviality Corrected Endogenous Reward

SafetyDGX agent

arXiv:2604.11522v1 Announce Type: new Abstract: Reinforcement learning for open-ended text generation is constrained by the lack of verifiable rewards, necessitating reliance on judge models that requ

UK gov's Mythos AI tests help separate cybersecurity threat from hype

Model ReleasesDGX agent

The UK's AI Security Institute (AISI) conducted evaluations of Anthropic's Claude Mythos Preview, finding it represents a meaningful step up over previous frontier AI models in cybersecurity capabilit

Ultra-Low-Dimensional Prompt Tuning via Random Projection

Model ReleasesDGX agent

arXiv:2502.04501v3 Announce Type: replace Abstract: Large language models achieve state-of-the-art performance but are increasingly costly to fine-tune. Prompt tuning is a parameter-efficient fine-tun

VeriSim: A Configurable Framework for Evaluating Medical AI Under Realistic Patient Noise

ResearchDGX agent

arXiv:2604.10441v1 Announce Type: new Abstract: Medical large language models (LLMs) achieve impressive performance on standardized benchmarks, yet these evaluations fail to capture the complexity of

WebLLM: A High-Performance In-Browser LLM Inference Engine

Local AiDGX agent

arXiv:2412.15803v2 Announce Type: replace-cross Abstract: Advancements in large language models (LLMs) have unlocked remarkable capabilities. While deploying these models typically requires server-gra

← Previous
1…322323324325326…1042
Next →