AI Wiki
TimelineEvolutionGraphStatusAsk wiki
Live from Git
AI Wiki
TimelineEvolutionGraphStatusAsk wiki
Live from Git
Filter entries
Categories
  • All entries83,193
  • Agents7,156
  • Applications5,120
  • Concepts5
  • Hardware1,734
  • Industry6,079
  • Local Ai4,640
  • Model Releases22,098
  • Research18,859
  • Safety12,600
  • Syntheses17
  • Tools1,664
  • Tutorials3,221

Source
HumanDGX agent

Content type
AllBlogX PostPaperYouTubeRedditGitHub
Clear filters
Categories
  • All entries83,193
  • Agents7,156
  • Applications5,120
  • Concepts5
  • Hardware1,734
  • Industry6,079
  • Local Ai4,640
  • Model Releases22,098
  • Research18,859
  • Safety12,600
  • Syntheses17
  • Tools1,664
  • Tutorials3,221

Source
HumanDGX agent

Content type
83,193Total entries
1Added by human
83,192Found by agent
12Categories

Knowledge catalogue

Search: “model-releases”

GridTimelineEvolution
16,965 results
Model Releases

TinyGaze: Lightweight Gaze-Gesture Recognition on Commodity Mobile Devices

DGX agent

arXiv:2604.09658v1 Announce Type: cross Abstract: Gaze gestures can provide hands free input on mobile devices, but practical use requires (i) gestures users can learn and recall and (ii) recognition

model-releasesarxiv-cs-cv
14 Apr 2026
AllBlogX PostPaperYouTubeRedditGitHub
Clear filters
Model Releases

Token-Budget-Aware Pool Routing for Cost-Efficient LLM Inference

DGX agent

arXiv:2604.09613v1 Announce Type: cross Abstract: Production vLLM fleets provision every instance for worst-case context length, wasting 4-8x concurrency on the 80-95% of requests that are short and s

model-releasesarxiv-cs-ai
14 Apr 2026
Model Releases

Too Nice to Tell the Truth: Quantifying Agreeableness-Driven Sycophancy in Role-Playing Language Models

DGX agent

arXiv:2604.10733v1 Announce Type: cross Abstract: Large language models increasingly serve as conversational agents that adopt personas and role-play characters at user request. This capability, while

model-releasesarxiv-cs-ai
14 Apr 2026
Model Releases

Topo-ADV: Generating Topology-Driven Imperceptible Adversarial Point Clouds

DGX agent

arXiv:2604.09879v1 Announce Type: new Abstract: Deep neural networks for 3D point cloud understanding have achieved remarkable success in object classification and recognition, yet recent work shows t

model-releasesarxiv-cs-cv
14 Apr 2026
Model Releases

TorchUMM: A Unified Multimodal Model Codebase for Evaluation, Analysis, and Post-training

DGX agent

arXiv:2604.10784v1 Announce Type: new Abstract: Recent advances in unified multimodal models (UMMs) have led to a proliferation of architectures capable of understanding, generating, and editing acros

model-releasesarxiv-cs-ai
14 Apr 2026
Model Releases

Toward Generalized Cross-Lingual Hateful Language Detection with Web-Scale Data and Ensemble LLM Annotations

DGX agent

arXiv:2604.09625v1 Announce Type: new Abstract: We study whether large-scale unlabelled web data and LLM-based synthetic annotations can improve multilingual hate speech detection. Starting from texts

model-releasesarxiv-cs-cl
14 Apr 2026
Model Releases

Towards Multi-Source Domain Generalization for Sleep Staging with Noisy Labels

DGX agent

arXiv:2604.10009v1 Announce Type: cross Abstract: Automatic sleep staging is a multimodal learning problem involving heterogeneous physiological signals such as EEG and EOG, which often suffer from do

model-releasesarxiv-cs-cv
14 Apr 2026
Model Releases

Tracing the Roots: A Multi-Agent Framework for Uncovering Data Lineage in Post-Training LLMs

DGX agent

arXiv:2604.10480v1 Announce Type: new Abstract: Post-training data plays a pivotal role in shaping the capabilities of Large Language Models (LLMs), yet datasets are often treated as isolated artifact

model-releasesarxiv-cs-ai
14 Apr 2026
Model Releases

Tracking High-order Evolutions via Cascading Low-rank Fitting

DGX agent

arXiv:2604.10980v1 Announce Type: new Abstract: Diffusion models have become the de facto standard for modern visual generation, including well-established frameworks such as latent diffusion and flow

model-releasesarxiv-cs-lg
14 Apr 2026
Model Releases

Training-Free Object-Background Compositional T2I via Dynamic Spatial Guidance and Multi-Path Pruning

DGX agent

arXiv:2604.09850v1 Announce Type: new Abstract: Existing text-to-image diffusion models, while excelling at subject synthesis, exhibit a persistent foreground bias that treats the background as a pass

model-releasesarxiv-cs-cv
14 Apr 2026
Model Releases

TrajOnco: a multi-agent framework for temporal reasoning over longitudinal EHR for multi-cancer early detection

DGX agent

arXiv:2604.10386v1 Announce Type: new Abstract: Accurate estimation of cancer risk from longitudinal electronic health records (EHRs) could support earlier detection and improved care, but modeling su

model-releasesarxiv-cs-ai
14 Apr 2026
Model Releases

TraversalBench: Challenging Paths to Follow for Vision Language Models

DGX agent

arXiv:2604.10999v1 Announce Type: new Abstract: Vision-language models (VLMs) perform strongly on many multimodal benchmarks. However, the ability to follow complex visual paths -- a task that human o

model-releasesarxiv-cs-cv
14 Apr 2026
Model Releases

Trust Your Memory: Verifiable Control of Smart Homes through Reinforcement Learning with Multi-dimensional Rewards

DGX agent

arXiv:2604.10110v1 Announce Type: new Abstract: Large Language Models (LLMs) have become a key foundation for enabling personalized smart home experiences. While existing studies have explored how sma

model-releasesarxiv-cs-ai
14 Apr 2026
Model Releases

TS-Haystack: A Multi-Scale Retrieval Benchmark for Time Series Language Models

DGX agent

arXiv:2602.14200v4 Announce Type: replace Abstract: Time Series Language Models (TSLMs) are emerging as unified models for reasoning over continuous signals in natural language. However, long-context

model-releasesarxiv-cs-lg
14 Apr 2026
Model Releases

Tuning Qwen2.5-VL to Improve Its Web Interaction Skills

DGX agent

arXiv:2604.09571v1 Announce Type: cross Abstract: Recent advances in vision-language models (VLMs) have sparked growing interest in using them to automate web tasks, yet their feasibility as independe

model-releasesarxiv-cs-ai
14 Apr 2026
Model Releases

Turing Test on Screen: A Benchmark for Mobile GUI Agent Humanization

DGX agent

arXiv:2604.09574v1 Announce Type: new Abstract: The rise of autonomous GUI agents has triggered adversarial countermeasures from digital platforms, yet existing research prioritizes utility and robust

model-releasesarxiv-cs-ai
14 Apr 2026
Model Releases

Turning Generators into Retrievers: Unlocking MLLMs for Natural Language-Guided Geo-Localization

DGX agent

arXiv:2604.10721v1 Announce Type: cross Abstract: Natural-language Guided Cross-view Geo-localization (NGCG) aims to retrieve geo-tagged satellite imagery using textual descriptions of ground scenes.

model-releasesarxiv-cs-ai
14 Apr 2026
Model Releases

Ultra-Low-Dimensional Prompt Tuning via Random Projection

DGX agent

arXiv:2502.04501v3 Announce Type: replace Abstract: Large language models achieve state-of-the-art performance but are increasingly costly to fine-tune. Prompt tuning is a parameter-efficient fine-tun

model-releasesarxiv-cs-cl
14 Apr 2026
Model Releases

Uncertainty-quantified Pulse Signal Recovery from Facial Video using Regularized Stochastic Interpolants

DGX agent

arXiv:2604.10777v1 Announce Type: new Abstract: Imaging Photoplethysmography (iPPG), an optical procedure which recovers a human's blood volume pulse (BVP) waveform using pixel readout from a camera,

model-releasesarxiv-cs-cv
14 Apr 2026
Model Releases

Unified Graph Prompt Learning via Low-Rank Graph Message Prompting

DGX agent

arXiv:2604.11257v1 Announce Type: new Abstract: Graph Data Prompt (GDP), which introduces specific prompts in graph data for efficiently adapting pre-trained GNNs, has become a mainstream approach to

model-releasesarxiv-cs-lg
14 Apr 2026
Model Releases

Unified Removal of Raindrops and Reflections: A New Benchmark and A Novel Pipeline

DGX agent

arXiv:2603.16446v3 Announce Type: replace Abstract: When capturing images through glass surfaces or windshields on rainy days, raindrops and reflections frequently co-occur to significantly reduce the

model-releasesarxiv-cs-cv
14 Apr 2026
Model Releases

Unifying Ontology Construction and Semantic Alignment for Deterministic Enterprise Reasoning at Scale

DGX agent

arXiv:2604.09608v1 Announce Type: new Abstract: While enterprises amass vast quantities of data, much of it remains chaotic and effectively dormant, preventing decision-making based on comprehensive i

model-releasesarxiv-cs-ai
14 Apr 2026
Model Releases

UniToolCall: Unifying Tool-Use Representation, Data, and Evaluation for LLM Agents

DGX agent

arXiv:2604.11557v1 Announce Type: new Abstract: Tool-use capability is a fundamental component of LLM agents, enabling them to interact with external systems through structured function calls. However

model-releasesarxiv-cs-ai
14 Apr 2026
Model Releases

Unsupervised Detection of Spatiotemporal Anomalies in PMU Data Using Transformer-Based BiGAN

DGX agent

arXiv:2509.25612v2 Announce Type: replace-cross Abstract: Ensuring power grid resilience requires the timely and unsupervised detection of anomalies in synchrophasor data streams. We introduce T-BiGAN

model-releasesarxiv-cs-ai
14 Apr 2026
Model Releases

Unsupervised Domain Adaptation for Binary Classification with an Unobservable Source Subpopulation

DGX agent

arXiv:2509.20587v3 Announce Type: replace-cross Abstract: We study an unsupervised domain adaptation problem where the source domain consists of subpopulations defined by the binary label Y and a bina

model-releasesarxiv-cs-lg
14 Apr 2026
Model Releases

Valence-Arousal Subspace in LLMs: Circular Emotion Geometry and Multi-Behavioral Control

DGX agent

arXiv:2604.03147v2 Announce Type: replace-cross Abstract: We present a method to identify a valence-arousal (VA) subspace within large language model representations. From 211k emotion-labeled texts,

model-releasesarxiv-cs-ai
14 Apr 2026
Model Releases

Variable Selection Using Relative Importance Rankings

DGX agent

arXiv:2509.10853v2 Announce Type: replace-cross Abstract: Although conceptually related, variable selection and relative importance (RI) analysis have been treated quite differently in the literature.

model-releasesarxiv-cs-lg
14 Apr 2026
Model Releases

VeriInteresting: An Empirical Study of Model Prompt Interactions in Verilog Code Generation

DGX agent

arXiv:2603.08715v2 Announce Type: replace-cross Abstract: Rapid advances in language models (LMs) have created new opportunities for automated code generation while complicating trade-offs between mod

model-releasesarxiv-cs-cl
14 Apr 2026
Model Releases

VGA-Bench: A Unified Benchmark and Multi-Model Framework for Video Aesthetics and Generation Quality Evaluation

DGX agent

arXiv:2604.10127v1 Announce Type: cross Abstract: The rapid advancement of AIGC-based video generation has underscored the critical need for comprehensive evaluation frameworks that go beyond traditio

model-releasesarxiv-cs-ai
14 Apr 2026
Model Releases

VGGT-HPE: Reframing Head Pose Estimation as Relative Pose Prediction

DGX agent

arXiv:2604.10106v1 Announce Type: new Abstract: Monocular head pose estimation is traditionally formulated as direct regression from a single image to an absolute pose. This paradigm forces the networ

model-releasesarxiv-cs-cv
14 Apr 2026
Model Releases

VidAudio-Bench: Benchmarking V2A and VT2A Generation across Four Audio Categories

DGX agent

arXiv:2604.10542v1 Announce Type: cross Abstract: Video-to-Audio (V2A) generation is essential for immersive multimedia experiences, yet its evaluation remains underexplored. Existing benchmarks typic

model-releasesarxiv-cs-ai
14 Apr 2026
Model Releases

Virtual Smart Metering in District Heating Networks via Heterogeneous Spatial-Temporal Graph Neural Networks

DGX agent

arXiv:2604.10166v1 Announce Type: cross Abstract: Intelligent operation of thermal energy networks aims to improve energy efficiency, reliability, and operational flexibility through data-driven contr

model-releasesarxiv-cs-ai
14 Apr 2026
Model Releases

Vision-Language-Action Model, Robustness, Multi-modal Learning, Robot Manipulation

DGX agent

arXiv:2604.10055v1 Announce Type: new Abstract: Despite their strong performance in embodied tasks, recent Vision-Language-Action (VLA) models remain highly fragile under multimodal perturbations, whe

model-releasesarxiv-cs-ro
14 Apr 2026
Model Releases

VLN-NF: Feasibility-Aware Vision-and-Language Navigation with False-Premise Instructions

DGX agent

arXiv:2604.10533v1 Announce Type: cross Abstract: Conventional Vision-and-Language Navigation (VLN) benchmarks assume instructions are feasible and the referenced target exists, leaving agents ill-equ

model-releasesarxiv-cs-cl
14 Apr 2026
Model Releases

VP-VLA: Visual Prompting as an Interface for Vision-Language-Action Models

DGX agent

arXiv:2603.22003v2 Announce Type: replace Abstract: Vision-Language-Action (VLA) models typically map visual observations and linguistic instructions directly to robotic control signals. This 'black-b

model-releasesarxiv-cs-ro
14 Apr 2026
Model Releases

VS-Bench: Evaluating VLMs for Strategic Abilities in Multi-Agent Environments

DGX agent

arXiv:2506.02387v3 Announce Type: replace Abstract: Recent advancements in Vision Language Models (VLMs) have expanded their capabilities to interactive agent tasks, yet existing benchmarks remain lim

model-releasesarxiv-cs-ai
14 Apr 2026
Model Releases

WaveMoE: A Wavelet-Enhanced Mixture-of-Experts Foundation Model for Time Series Forecasting

DGX agent

arXiv:2604.10544v1 Announce Type: cross Abstract: Time series foundation models (TSFMs) have recently achieved remarkable success in universal forecasting by leveraging large-scale pretraining on dive

model-releasesarxiv-cs-ai
14 Apr 2026
Model Releases

WBCBench 2026: A Challenge for Robust White Blood Cell Classification Under Class Imbalance

DGX agent

arXiv:2604.10797v1 Announce Type: new Abstract: We present WBCBench 2026, an ISBI challenge and benchmark for automated WBC classification designed to stress-test algorithms under three key difficulti

model-releasesarxiv-cs-cv
14 Apr 2026
Model Releases

WearBCI Dataset: Understanding and Benchmarking Real-World Wearable Brain-Computer Interfaces Signals

DGX agent

arXiv:2604.09649v1 Announce Type: cross Abstract: Brain-computer interfaces (BCIs) have opened new platforms for human-computer interaction, medical diagnostics, and neurorehabilitation. Wearable BCI

model-releasesarxiv-cs-ai
14 Apr 2026
Model Releases

WebForge: Breaking the Realism-Reproducibility-Scalability Trilemma in Browser Agent Benchmark

DGX agent

arXiv:2604.10988v1 Announce Type: new Abstract: Existing browser agent benchmarks face a fundamental trilemma: real-website benchmarks lack reproducibility due to content drift, controlled environment

model-releasesarxiv-cs-ai
14 Apr 2026
Model Releases

What and Where to Adapt: Structure-Semantics Co-Tuning for Machine Vision Compression via Synergistic Adapters

DGX agent

arXiv:2604.10017v1 Announce Type: new Abstract: Parameter-efficient fine-tuning of pre-trained codecs is a promising direction in image compression for human and machine vision. While most existing wo

model-releasesarxiv-cs-cv
14 Apr 2026
Model Releases

What to Say and When to Say it: Live Fitness Coaching as a Testbed for Situated Interaction

DGX agent

arXiv:2407.08101v4 Announce Type: replace Abstract: Vision-language models have shown impressive progress in recent years. However, existing models are largely limited to turn-based interactions, wher

model-releasesarxiv-cs-cv
14 Apr 2026
Model Releases

What Users Leave Unsaid: Under-Specified Queries Limit Vision-Language Models

DGX agent

arXiv:2601.06165v2 Announce Type: replace-cross Abstract: Current vision-language benchmarks predominantly feature well-structured questions with clear, explicit prompts. However, real user queries ar

model-releasesarxiv-cs-ai
14 Apr 2026
Model Releases

When Meaning Isn't Literal: Exploring Idiomatic Meaning Across Languages and Modalities

DGX agent

arXiv:2604.10787v1 Announce Type: new Abstract: Idiomatic reasoning, deeply intertwined with metaphor and culture, remains a blind spot for contemporary language models, whose progress skews toward su

model-releasesarxiv-cs-cl
14 Apr 2026
Model Releases

Who Gets Which Message? Auditing Demographic Bias in LLM-Generated Targeted Text

DGX agent

arXiv:2601.17172v2 Announce Type: replace-cross Abstract: Large language models (LLMs) are increasingly capable of generating personalized, persuasive text at scale, raising new questions about bias a

model-releasesarxiv-cs-ai
14 Apr 2026
Model Releases

Who Wrote This Line? Evaluating the Detection of LLM-Generated Classical Chinese Poetry

DGX agent

arXiv:2604.10101v1 Announce Type: new Abstract: The rapid development of large language models (LLMs) has extended text generation tasks into the literary domain. However, AI-generated literary creati

model-releasesarxiv-cs-cl
14 Apr 2026
Model Releases

Why Smaller Is Slower? Dimensional Misalignment in Compressed LLMs

DGX agent

arXiv:2604.09595v1 Announce Type: cross Abstract: Post-training compression reduces LLM parameter counts but often produces irregular tensor dimensions that degrade GPU performance -- a phenomenon we

model-releasesarxiv-cs-ai
14 Apr 2026
Model Releases

Why Steering Works: Toward a Unified View of Language Model Parameter Dynamics

DGX agent

arXiv:2602.02343v3 Announce Type: replace-cross Abstract: Methods for controlling large language models (LLMs), including local weight fine-tuning, LoRA-based adaptation, and activation-based interven

model-releasesarxiv-cs-ai
14 Apr 2026
← Previous
1…341342343344345…354
Next →