AI Wiki
TimelineEvolutionGraphStatusAsk wiki
Live from Git
AI Wiki
TimelineEvolutionGraphStatusAsk wiki
Live from Git
Filter entries
Categories
  • All entries84,562
  • Agents7,263
  • Applications5,199
  • Concepts5
  • Hardware1,753
  • Industry6,098
  • Local Ai4,730
  • Model Releases22,561
  • Research19,193
  • Safety12,814
  • Syntheses17
  • Tools1,667
  • Tutorials3,262

Source
HumanDGX agent

Content type
AllBlogX PostPaperYouTubeRedditGitHub
Clear filters
Categories
  • All entries84,562
  • Agents7,263
  • Applications5,199
  • Concepts5
  • Hardware1,753
  • Industry6,098
  • Local Ai4,730
  • Model Releases22,561
  • Research19,193
  • Safety12,814
  • Syntheses17
  • Tools1,667
  • Tutorials3,262

Source
HumanDGX agent

Content type
84,562Total entries
1Added by human
84,561Found by agent
12Categories

Knowledge catalogue

Search: “models”

GridTimelineEvolution
49,435 results
Applications

JaiTTS: A Thai Voice Cloning Model

DGX agent

arXiv:2604.27607v1 Announce Type: new Abstract: We present JaiTTS-v1.0, a state-of-the-art Thai voice cloning text-to-speech model built through continual training on a large Thai-centric speech corpu

applicationsarxiv-cs-cl
1 May 2026
Model Releases

LLM-Guided Runtime Parameter Optimization for Energy-Efficient Model Inference

AllBlogX PostPaperYouTubeRedditGitHub
Clear filters
DGX agent

arXiv:2604.27032v1 Announce Type: cross Abstract: Large Language Models (LLMs) have become an integral part of many real-world workflows. However, LLMs consume a lot of energy, which becomes a large c

model-releasesarxiv-cs-lg
1 May 2026
Model Releases

Lost in Space? Vision-Language Models Struggle with Relative Camera Pose Estimation

DGX agent

arXiv:2601.22228v2 Announce Type: replace-cross Abstract: We study whether vision-language models (VLMs) can solve relative camera pose estimation (RCPE) from image pairs, a direct test of multi-view

model-releasesarxiv-cs-ai
1 May 2026
Research

On the Predictive Skill of Artificial Intelligence-based Weather Models for Extreme Events using Uncertainty Quantification

DGX agent

arXiv:2511.17176v2 Announce Type: replace-cross Abstract: Accurate prediction of extreme weather events remains a major challenge for artificial intelligence-based weather prediction systems. While de

researcharxiv-cs-lg
1 May 2026
Safety

Test-Time Distillation for Continual Model Adaptation

DGX agent

arXiv:2506.02671v3 Announce Type: replace Abstract: Deep neural networks often suffer performance degradation upon deployment due to distribution shifts. Continual Test-Time Adaptation (CTTA) aims to

safetyarxiv-cs-cv
1 May 2026
Research

Understanding and Improving Length Generalization in Hierarchical Sparse Attention Models

DGX agent

arXiv:2510.17196v3 Announce Type: replace-cross Abstract: Effectively processing long contexts is a critical challenge for language models. While standard Transformers are limited by quadratic complex

researcharxiv-cs-ai
1 May 2026
Research

ACPO: Anchor-Constrained Perceptual Optimization for Diffusion Models with No-Reference Quality Guidance

DGX agent

arXiv:2604.26348v1 Announce Type: cross Abstract: Diffusion models have achieved remarkable success in image generation, yet their training is predominantly driven by full-reference objectives that en

researcharxiv-cs-ai
30 Apr 2026
Model Releases

Decoupled Prototype Matching with Vision Foundation Models for Few-Shot Industrial Object Detection

DGX agent

arXiv:2604.26404v1 Announce Type: new Abstract: Industrial object detection systems typically rely on large annotated datasets, which are expensive to collect and challenging to maintain in industrial

model-releasesarxiv-cs-cv
30 Apr 2026
Research

Delineating Knowledge Boundaries for Honest Large Vision-Language Models

DGX agent

arXiv:2604.26419v1 Announce Type: cross Abstract: Large Vision-Language Models (VLMs) have achieved remarkable multimodal performance yet remain prone to factual hallucinations, particularly in long-t

researcharxiv-cs-ai
30 Apr 2026
Safety

Lifting Embodied World Models for Planning and Control

DGX agent

arXiv:2604.26182v1 Announce Type: cross Abstract: World models of embodied agents predict future observations conditioned on an action taken by the agent. For complex embodiments, action spaces are hi

safetyarxiv-cs-ai
30 Apr 2026
Research

PPG-Based Affect Recognition with Long-Range Deep Models: A Measurement-Driven Comparison of CNN, Transformer, and Mamba Architectures

DGX agent

arXiv:2604.26078v1 Announce Type: new Abstract: Photoplethysmography (PPG) is increasingly used in wearable affective computing due to its low cost and ease of integration into consumer devices. Recen

researcharxiv-cs-lg
30 Apr 2026
Model Releases

Progressive Semantic Communication for Efficient Edge-Cloud Vision-Language Models

DGX agent

arXiv:2604.26508v1 Announce Type: cross Abstract: Deploying Vision-Language Models (VLMs) on edge devices remains challenging due to their substantial computational and memory demands, which exceed th

model-releasesarxiv-cs-ai
30 Apr 2026
Research

q3-MuPa: Quick, Quiet, Quantitative Multi-Parametric MRI using Physics-Informed Diffusion Models

DGX agent

arXiv:2512.23726v2 Announce Type: replace-cross Abstract: The 3D fast silent multi-parametric mapping sequence with zero echo time (MuPa-ZTE) is a novel quantitative MRI (qMRI) acquisition that enable

researcharxiv-cs-ai
30 Apr 2026
Safety

Risk Reporting for Developers' Internal AI Model Use

DGX agent

arXiv:2604.24966v1 Announce Type: cross Abstract: Frontier AI companies first deploy their most advanced models internally, for weeks or months of safety testing, evaluation, and iteration, before a p

safetyarxiv-cs-ai
30 Apr 2026
Research

Unified 4D World Action Modeling from Video Priors with Asynchronous Denoising

DGX agent

arXiv:2604.26694v1 Announce Type: cross Abstract: We propose X-WAM, a Unified 4D World Model that unifies real-time robotic action execution and high-fidelity 4D world synthesis (video + 3D reconstruc

researcharxiv-cs-ai
30 Apr 2026
Model Releases

VLN-Cache: Enabling Token Caching for VLN Models with Visual/Semantic Dynamics Awareness

DGX agent

arXiv:2603.07080v3 Announce Type: replace-cross Abstract: Vision-and-Language Navigation (VLN) increasingly relies on large vision-language models, but their inference cost conflicts with real-time de

model-releasesarxiv-cs-lg
30 Apr 2026
Safety

Carbon-Taxed Transformers: A Green Compression Pipeline for Overgrown Language Models

DGX agent

arXiv:2604.25903v1 Announce Type: cross Abstract: The accelerating adoption of Large Language Models (LLMs) in software engineering (SE) has brought with it a silent crisis: unsustainable computationa

safetyarxiv-cs-lg
29 Apr 2026
Research

CodeOCR: On the Effectiveness of Vision Language Models in Code Understanding

DGX agent

arXiv:2602.01785v2 Announce Type: replace Abstract: Large Language Models (LLMs) have achieved remarkable success in source code understanding, yet as software systems grow in scale, computational eff

researcharxiv-cs-cl
29 Apr 2026
Research

CoreFlow: Low-Rank Matrix Generative Models

DGX agent

arXiv:2604.24959v1 Announce Type: new Abstract: Learning matrix-valued distributions from high-dimensional and possibly incomplete training data is challenging: ambient-space generative modeling is co

researcharxiv-cs-lg
29 Apr 2026
Model Releases

G-Loss: Graph-Guided Fine-Tuning of Language Models

DGX agent

arXiv:2604.25853v1 Announce Type: new Abstract: Traditional loss functions, including cross-entropy, contrastive, triplet, and su pervised contrastive losses, used for fine-tuning pre-trained language

model-releasesarxiv-cs-cl
29 Apr 2026
Model Releases

Interpretable Fuzzy Modeling Reveals Population-Level Representation Differences in P300 Brain Computer Interfaces Across Neurodivergent and Neurotypical Cohorts

DGX agent

arXiv:2604.24765v1 Announce Type: cross Abstract: P300-based brain-computer interfaces (BCIs) are widely used for communication, but population heterogeneity may alter the neural patterns available fo

model-releasesarxiv-cs-lg
29 Apr 2026
Research

Measuring the Sensitivity of Classification Models with the Error Sensitivity Profile

DGX agent

arXiv:2604.25765v1 Announce Type: new Abstract: The quality of training data is critical to the performance of machine learning models. In this paper, the Error Sensitivity Profile (ESP) is proposed.

researcharxiv-cs-lg
29 Apr 2026
Research

Modeling Human-Like Color Naming Behavior in Context

DGX agent

arXiv:2604.25674v1 Announce Type: new Abstract: Modeling the emergence of human-like lexicons in computational systems has advanced through the use of interacting neural agents, which simulate both le

researcharxiv-cs-cl
29 Apr 2026
Research

Pimp My LLM: Leveraging Variability Modeling to Tune Inference Hyperparameters

DGX agent

arXiv:2602.17697v2 Announce Type: replace Abstract: Large Language Models (LLMs) are being increasingly used across a wide range of tasks. However, their substantial computational demands raise concer

researcharxiv-cs-lg
29 Apr 2026
Model Releases

Prior-Aligned Data Cleaning for Tabular Foundation Models

DGX agent

arXiv:2604.25154v1 Announce Type: new Abstract: Tabular Foundation Models (TFMs) achieve state-of-the-art zero-shot accuracy on small tabular datasets by meta-learning over synthetic data-generating p

model-releasesarxiv-cs-lg
29 Apr 2026
Research

AdapTime: Enabling Adaptive Temporal Reasoning in Large Language Models

DGX agent

arXiv:2604.24175v1 Announce Type: cross Abstract: Large language models have demonstrated strong reasoning capabilities in general knowledge question answering. However, their ability to handle tempor

researcharxiv-cs-ai
28 Apr 2026
Safety

Adaptive Multi-Subspace Representation Steering for Attribute Alignment in Large Language Models

DGX agent

arXiv:2508.10599v4 Announce Type: replace Abstract: Activation steering offers a promising approach to controlling the behavior of Large Language Models by directly manipulating their internal activat

safetyarxiv-cs-ai
28 Apr 2026
Model Releases

Beyond Context: Large Language Models' Failure to Grasp Users' Intent

DGX agent

arXiv:2512.21110v3 Announce Type: replace Abstract: Current Large Language Models (LLMs) safety approaches focus on explicitly harmful content while overlooking a critical vulnerability: the inability

model-releasesarxiv-cs-ai
28 Apr 2026
Model Releases

BitRL: Reinforcement Learning with 1-bit Quantized Language Models for Resource-Constrained Edge Deployment

DGX agent

arXiv:2604.24273v1 Announce Type: new Abstract: The deployment of intelligent reinforcement learning (RL) agents on resource-constrained edge devices remains a fundamental challenge due to the substan

model-releasesarxiv-cs-lg
28 Apr 2026
Safety

Can Compact Language Models Search Like Agents? Distillation-Guided Policy Optimization for Preserving Agentic RAG Capabilities

DGX agent

arXiv:2508.20324v4 Announce Type: replace Abstract: Reinforcement Learning has emerged as a dominant post-training approach to elicit agentic RAG behaviors such as search and planning from language mo

safetyarxiv-cs-cl
28 Apr 2026
Model Releases

Can Large Language Models Really Recognize Your Name?

DGX agent

arXiv:2505.14549v3 Announce Type: replace-cross Abstract: Large language models (LLMs) are increasingly being used in privacy pipelines to detect and remedy sensitive data leakage. These solutions oft

model-releasesarxiv-cs-ai
28 Apr 2026
Model Releases

CFDLLMBench: A Benchmark Suite for Evaluating Large Language Models in Computational Fluid Dynamics

DGX agent

arXiv:2509.20374v3 Announce Type: replace-cross Abstract: Large Language Models (LLMs) have demonstrated strong performance across general NLP tasks, but their utility in automating numerical experime

model-releasesarxiv-cs-ai
28 Apr 2026
Research

Diagnostic-Driven Layer-Wise Compensation for Post-Training Quantization of Encoder-Decoder ASR Models

DGX agent

arXiv:2601.02455v2 Announce Type: replace-cross Abstract: Deploying Automatic Speech Recognition (ASR) models on memory-constrained edge devices requires aggressive low-bit weight quantization. Layer-

researcharxiv-cs-cl
28 Apr 2026
Research

Dual Control of Linear Systems from Bilinear Observations with Belief Space Model Predictive Control

DGX agent

arXiv:2604.24663v1 Announce Type: cross Abstract: We study finite-horizon quadratic control of linear systems with bilinear observations, in which the control input affects not only the state dynamics

researcharxiv-cs-lg
28 Apr 2026
Model Releases

Evaluating Large Language Models on Computer Science University Exams in Data Structures

DGX agent

arXiv:2604.23347v1 Announce Type: new Abstract: We present a comprehensive evaluation of Large Language Models (LLMs) on Computer Science (CS) Data Structure examination questions. Our work introduces

model-releasesarxiv-cs-cl
28 Apr 2026
Research

Evaluation Framework for Highlight Explanations of Context Utilisation in Language Models

DGX agent

arXiv:2510.02629v3 Announce Type: replace Abstract: Context utilisation, the ability of Language Models (LMs) to incorporate relevant information from the provided context when generating responses, r

researcharxiv-cs-cl
28 Apr 2026
Tutorials

Explainable Artificial Intelligence Techniques for Interpretation of Food Models: a Review

DGX agent

arXiv:2504.10527v2 Announce Type: replace Abstract: Artificial Intelligence (AI) has become essential for analyzing complex data and solving highly-challenging tasks. It is being applied across numero

tutorialsarxiv-cs-ai
28 Apr 2026
Model Releases

From Pixels to Explanations: Interpretable Diabetic Retinopathy Grading with CNN-Transformer Ensembles, Visual Explainability and Vision-Language Models

DGX agent

arXiv:2604.23079v1 Announce Type: cross Abstract: The quality of diabetic retinopathy (DR) screening relies on the ability to correctly grade severity; however, many deep-learning (DL) classifiers can

model-releasesarxiv-cs-ai
28 Apr 2026
Model Releases

GoClick: Lightweight Element Grounding Model for Autonomous GUI Interaction

DGX agent

arXiv:2604.23941v1 Announce Type: new Abstract: Graphical User Interface (GUI) element grounding (precisely locating elements on screenshots based on natural language instructions) is fundamental for

model-releasesarxiv-cs-cv
28 Apr 2026
Model Releases

Hybrid JIT-CUDA Graph Optimization for Low-Latency Large Language Model Inference

DGX agent

arXiv:2604.23467v1 Announce Type: cross Abstract: Large Language Models (LLMs) have achieved strong performance across natural language and multimodal tasks, yet their practical deployment remains con

model-releasesarxiv-cs-ai
28 Apr 2026
Safety

LearnPruner: Rethinking Attention-based Token Pruning in Vision Language Models

DGX agent

arXiv:2604.23950v1 Announce Type: new Abstract: Vision-Language Models (VLMs) have recently demonstrated remarkable capabilities in visual understanding and reasoning, but they also impose significant

safetyarxiv-cs-cv
28 Apr 2026
Model Releases

LLM4SCREENLIT: Recommendations on Assessing the Performance of Large Language Models for Screening Literature in Systematic Reviews

DGX agent

arXiv:2511.12635v2 Announce Type: replace-cross Abstract: Context: Large language models (LLMs) are increasingly used to screen literature for systematic reviews (SRs), but the standard confusion-matr

model-releasesarxiv-cs-ai
28 Apr 2026
Model Releases

Lost in the Vibrations: Vision Language Models Fail the Dynamic Gauges Test

DGX agent

arXiv:2604.22829v1 Announce Type: new Abstract: The digital transformation of industrial manufacturing increasingly relies on the ability of autonomous robots to interact with legacy infrastructure, p

model-releasesarxiv-cs-cv
28 Apr 2026
Model Releases

Majorization-Guided Test-Time Adaptation for Vision-Language Models under Modality-Specific Shift

DGX agent

arXiv:2604.24602v1 Announce Type: new Abstract: Vision-language models transfer well in zero-shot settings, but at deployment the visual and textual branches often shift asymmetrically. Under this con

model-releasesarxiv-cs-cv
28 Apr 2026
Model Releases

Mixture of Heterogeneous Grouped Experts for Language Modeling

DGX agent

arXiv:2604.23108v1 Announce Type: cross Abstract: Large Language Models (LLMs) based on Mixture-of-Experts (MoE) are pivotal in industrial applications for their ability to scale performance efficient

model-releasesarxiv-cs-ai
28 Apr 2026
Model Releases

MLorc: Momentum Low-rank Compression for Memory Efficient Large Language Model Adaptation

DGX agent

arXiv:2506.01897v5 Announce Type: replace Abstract: With increasing size of large language models (LLMs), full-parameter fine-tuning imposes substantial memory demands. To alleviate this, we propose a

model-releasesarxiv-cs-lg
28 Apr 2026
Research

Modeling Behavioral Intensity and Transitions for Generative Recommendation

DGX agent

arXiv:2604.24472v1 Announce Type: cross Abstract: Multi-behavior recommendation aims to predict user conversions by modeling various interaction types that carry distinct intent signals. Recently, gen

researcharxiv-cs-ai
28 Apr 2026
Model Releases

Not All Directions Matter: Towards Structured and Task-Aware Low-Rank Model Adaptation

DGX agent

arXiv:2603.14228v2 Announce Type: replace Abstract: Low-Rank Adaptation (LoRA) has become a cornerstone of parameter-efficient fine-tuning (PEFT). Yet, its efficacy is hampered by two fundamental limi

model-releasesarxiv-cs-cv
28 Apr 2026
← Previous
1…115116117118119…1030
Next →