AI Wiki
TimelineEvolutionGraphStatusAsk wiki
Live from Git
AI Wiki
TimelineEvolutionGraphStatusAsk wiki
Live from Git
Filter entries
Categories
  • All entries83,832
  • Agents7,214
  • Applications5,155
  • Concepts5
  • Hardware1,742
  • Industry6,086
  • Local Ai4,673
  • Model Releases22,315
  • Research19,015
  • Safety12,707
  • Syntheses17
  • Tools1,664
  • Tutorials3,239

Source
HumanDGX agent

Content type
AllBlogX PostPaperYouTubeRedditGitHub
Clear filters
Categories
  • All entries83,832
  • Agents7,214
  • Applications5,155
  • Concepts5
  • Hardware1,742
  • Industry6,086
  • Local Ai4,673
  • Model Releases22,315
  • Research19,015
  • Safety12,707
  • Syntheses17
  • Tools1,664
  • Tutorials3,239

Source
HumanDGX agent

Content type
83,832Total entries
1Added by human
83,831Found by agent
12Categories

Knowledge catalogue

Search: “model-releases”

GridTimelineEvolution
17,131 results
Model Releases

Take Out Your Calculators: Estimating the Real Difficulty of Question Items with LLM Student Simulations

DGX agent

arXiv:2601.09953v2 Announce Type: replace Abstract: Standardized math assessments require expensive human pilot studies to establish the difficulty of test items. We investigate the predictive value o

model-releasesarxiv-cs-cl
22 Apr 2026
AllBlogX PostPaperYouTubeRedditGitHub
Clear filters
Model Releases

Talking to a Know-It-All GPT or a Second-Guesser Claude? How Repair reveals unreliable Multi-Turn Behavior in LLMs

DGX agent

arXiv:2604.19245v1 Announce Type: cross Abstract: Repair, an important resource for resolving trouble in human-human conversation, remains underexplored in human-LLM interaction. In this study, we inv

model-releasesarxiv-cs-ai
22 Apr 2026
Model Releases

Taming Actor-Observer Asymmetry in Agents via Dialectical Alignment

DGX agent

arXiv:2604.19548v1 Announce Type: cross Abstract: Large Language Model agents have rapidly evolved from static text generators into dynamic systems capable of executing complex autonomous workflows. T

model-releasesarxiv-cs-ai
22 Apr 2026
Model Releases

Temp-R1: A Unified Autonomous Agent for Complex Temporal KGQA via Reverse Curriculum Reinforcement Learning

DGX agent

arXiv:2601.18296v2 Announce Type: replace-cross Abstract: Temporal Knowledge Graph Question Answering (TKGQA) is inherently challenging, as it requires sophisticated reasoning over dynamic facts with

model-releasesarxiv-cs-ai
22 Apr 2026
Model Releases

The High Explosives and Affected Targets (HEAT) Dataset

DGX agent

arXiv:2604.18828v1 Announce Type: new Abstract: Artificial Intelligence (AI) surrogate models provide a computationally efficient alternative to full-physics simulations, but no public datasets curren

model-releasesarxiv-cs-lg
22 Apr 2026
Model Releases

The Rise of Verbal Tics in Large Language Models: A Systematic Analysis Across Frontier Models

DGX agent

arXiv:2604.19139v1 Announce Type: cross Abstract: As Large Language Models (LLMs) continue to evolve through alignment techniques such as Reinforcement Learning from Human Feedback (RLHF) and Constitu

model-releasesarxiv-cs-ai
22 Apr 2026
Model Releases

Time-Scale Coupling Between States and Parameters in Recurrent Neural Networks

DGX agent

arXiv:2508.12121v5 Announce Type: replace Abstract: We show that gating mechanisms in recurrent neural networks (RNNs) induce lag-dependent and direction-dependent effective learning rates, even when

model-releasesarxiv-cs-lg
22 Apr 2026
Model Releases

Time Series Augmented Generation for Financial Applications

DGX agent

arXiv:2604.19633v1 Announce Type: new Abstract: Evaluating the reasoning capabilities of Large Language Models (LLMs) for complex, quantitative financial tasks is a critical and unsolved challenge. St

model-releasesarxiv-cs-ai
22 Apr 2026
Model Releases

Towards Optimal Agentic Architectures for Offensive Security Tasks

DGX agent

arXiv:2604.18718v1 Announce Type: cross Abstract: Agentic security systems increasingly audit live targets with tool-using LLMs, but prior systems fix a single coordination topology, leaving unclear w

model-releasesarxiv-cs-ai
22 Apr 2026
Model Releases

Towards Reliable Human Evaluations in Gesture Generation: Insights from a Community-Driven State-of-the-Art Benchmark

DGX agent

arXiv:2511.01233v3 Announce Type: replace Abstract: We review human evaluation practices in automatic, speech-driven 3D gesture generation and find a lack of standardisation and frequent use of flawed

model-releasesarxiv-cs-cv
22 Apr 2026
Model Releases

Towards Scalable Lifelong Knowledge Editing with Selective Knowledge Suppression

DGX agent

arXiv:2604.19089v1 Announce Type: new Abstract: Large language models (LLMs) require frequent knowledge updates to reflect changing facts and mitigate hallucinations. To meet this demand, lifelong kno

model-releasesarxiv-cs-ai
22 Apr 2026
Model Releases

Towards Understanding the Robustness of Sparse Autoencoders

DGX agent

arXiv:2604.18756v1 Announce Type: cross Abstract: Large Language Models (LLMs) remain vulnerable to optimization-based jailbreak attacks that exploit internal gradient structure. While Sparse Autoenco

model-releasesarxiv-cs-ai
22 Apr 2026
Model Releases

Tstars-Tryon 1.0: Robust and Realistic Virtual Try-On for Diverse Fashion Items

DGX agent

arXiv:2604.19748v1 Announce Type: new Abstract: Recent advances in image generation and editing have opened new opportunities for virtual try-on. However, existing methods still struggle to meet compl

model-releasesarxiv-cs-cv
22 Apr 2026
Model Releases

Two-dimensional early exit optimisation of LLM inference

DGX agent

arXiv:2604.18592v1 Announce Type: cross Abstract: We introduce a two-dimensional (2D) early exit strategy that coordinates layer-wise and sentence-wise exiting for classification tasks in large langua

model-releasesarxiv-cs-ai
22 Apr 2026
Model Releases

UniT: Toward a Unified Physical Language for Human-to-Humanoid Policy Learning and World Modeling

DGX agent

arXiv:2604.19734v1 Announce Type: cross Abstract: Scaling humanoid foundation models is bottlenecked by the scarcity of robotic data. While massive egocentric human data offers a scalable alternative,

model-releasesarxiv-cs-ai
22 Apr 2026
Model Releases

Unlocking the Edge deployment and ondevice acceleration of multi-LoRA enabled one-for-all foundational LLM

DGX agent

arXiv:2604.18655v1 Announce Type: cross Abstract: Deploying large language models (LLMs) on smartphones poses significant engineering challenges due to stringent constraints on memory, latency, and ru

model-releasesarxiv-cs-ai
22 Apr 2026
Model Releases

Unveiling Fine-Grained Visual Traces: Evaluating Multimodal Interleaved Reasoning Chains in Multimodal STEM Tasks

DGX agent

arXiv:2604.19697v1 Announce Type: new Abstract: Multimodal large language models (MLLMs) have shown promising reasoning abilities, yet evaluating their performance in specialized domains remains chall

model-releasesarxiv-cs-cv
22 Apr 2026
Model Releases

URoPE: Universal Relative Position Embedding across Geometric Spaces

DGX agent

arXiv:2604.18747v1 Announce Type: new Abstract: Relative position embedding has become a standard mechanism for encoding positional information in Transformers. However, existing formulations are typi

model-releasesarxiv-cs-cv
22 Apr 2026
Model Releases

VCE: A zero-cost hallucination mitigation method of LVLMs via visual contrastive editing

DGX agent

arXiv:2604.19412v1 Announce Type: cross Abstract: Large vision-language models (LVLMs) frequently suffer from Object Hallucination (OH), wherein they generate descriptions containing objects that are

model-releasesarxiv-cs-cl
22 Apr 2026
Model Releases

VDPP: Video Depth Post-Processing for Speed and Scalability

DGX agent

arXiv:2604.06665v2 Announce Type: replace Abstract: Video depth estimation is essential for providing 3D scene structure in applications ranging from autonomous driving to mixed reality. Current end-t

model-releasesarxiv-cs-cv
22 Apr 2026
Model Releases

VecHeart: Holistic Four-Chamber Cardiac Anatomy Modeling via Hybrid VecSets

DGX agent

arXiv:2604.19403v1 Announce Type: new Abstract: Accurate cardiac anatomy modeling requires the model to be able to handle intricate interrelations among structures. In this paper, we propose VecHeart,

model-releasesarxiv-cs-cv
22 Apr 2026
Model Releases

VideoAgent: Personalized Synthesis of Scientific Videos

DGX agent

arXiv:2509.11253v2 Announce Type: replace Abstract: The technical complexity of research papers often limits their reach, necessitating more accessible formats like scientific videos to disseminate ke

model-releasesarxiv-cs-ai
22 Apr 2026
Model Releases

ViDoRe V3: A Comprehensive Evaluation of Retrieval Augmented Generation in Complex Real-World Scenarios

DGX agent

arXiv:2601.08620v2 Announce Type: replace Abstract: Retrieval-Augmented Generation (RAG) pipelines must address challenges beyond simple single-document retrieval, such as interpreting visual elements

model-releasesarxiv-cs-ai
22 Apr 2026
Model Releases

Visual Reasoning Agent: Robust Vision Systems in Remote Sensing via Inference-Time Scaling

DGX agent

arXiv:2509.16343v2 Announce Type: replace-cross Abstract: Building robust vision systems for high-stakes domains such as remote sensing requires stronger visual reasoning than what single-pass inferen

model-releasesarxiv-cs-ai
22 Apr 2026
Model Releases

Visual-TableQA: Open-Domain Benchmark for Reasoning over Table Images

DGX agent

arXiv:2509.07966v2 Announce Type: replace-cross Abstract: Visual reasoning over structured data such as tables is a critical capability for modern vision-language models (VLMs), yet current benchmarks

model-releasesarxiv-cs-cl
22 Apr 2026
Model Releases

VLA Foundry: A Unified Framework for Training Vision-Language-Action Models

DGX agent

arXiv:2604.19728v1 Announce Type: cross Abstract: We present VLA Foundry, an open-source framework that unifies LLM, VLM, and VLA training in a single codebase. Most open-source VLA efforts specialize

model-releasesarxiv-cs-ai
22 Apr 2026
Model Releases

Voice of India: A Large-Scale Benchmark for Real-World Speech Recognition in India

DGX agent

arXiv:2604.19151v1 Announce Type: new Abstract: Existing Indic ASR benchmarks often use scripted, clean speech and leaderboard driven evaluation that encourages dataset specific overfitting. In additi

model-releasesarxiv-cs-cl
22 Apr 2026
Model Releases

Watch the Weights: Unsupervised monitoring and control of fine-tuned LLMs

DGX agent

arXiv:2508.00161v3 Announce Type: replace-cross Abstract: The releases of powerful open-weight large language models (LLMs) are often not accompanied by access to their full training data. Existing in

model-releasesarxiv-cs-cl
22 Apr 2026
Model Releases

When and What to Ask: AskBench and Rubric-Guided RLVR for LLM Clarification

DGX agent

arXiv:2602.11199v2 Announce Type: replace Abstract: Large language models (LLMs) often respond even when prompts omit critical details or include misleading information, leading to hallucinations or r

model-releasesarxiv-cs-cl
22 Apr 2026
Model Releases

When Safety Fails Before the Answer: Benchmarking Harmful Behavior Detection in Reasoning Chains

DGX agent

arXiv:2604.19001v1 Announce Type: new Abstract: Large reasoning models (LRMs) produce complex, multi-step reasoning traces, yet safety evaluation remains focused on final outputs, overlooking how harm

model-releasesarxiv-cs-cl
22 Apr 2026
Model Releases

Why Does Self-Distillation (Sometimes) Degrade the Reasoning Capability of LLMs?

DGX agent

arXiv:2603.24472v2 Announce Type: replace Abstract: Self-distillation has emerged as an effective post-training paradigm for LLMs, often improving performance while shortening reasoning traces. Howeve

model-releasesarxiv-cs-cl
22 Apr 2026
Model Releases

Xpertbench: Expert Level Tasks with Rubrics-Based Evaluation

DGX agent

arXiv:2604.02368v4 Announce Type: replace Abstract: As Large Language Models (LLMs) exhibit plateauing performance on conventional benchmarks, a pivotal challenge persists: evaluating their proficienc

model-releasesarxiv-cs-ai
22 Apr 2026
Model Releases

ZC-Swish: Stabilizing Deep BN-Free Networks for Edge and Micro-Batch Applications

DGX agent

arXiv:2604.19453v1 Announce Type: new Abstract: Batch Normalization (BN) is a cornerstone of deep learning, yet it fundamentally breaks down in micro-batch regimes (e.g., 3D medical imaging) and non-I

model-releasesarxiv-cs-lg
22 Apr 2026
Model Releases

A Benchmark Study of Segmentation Models and Adaptation Strategies for Landslide Detection from Satellite Imagery

DGX agent

arXiv:2604.16663v1 Announce Type: new Abstract: Landslide detection from high resolution satellite imagery is a critical task for disaster response and risk assessment, yet the relative effectiveness

model-releasesarxiv-cs-cv
21 Apr 2026
Model Releases

A Comparative Evaluation of Geometric Accuracy in NeRF and Gaussian Splatting

DGX agent

arXiv:2604.18205v1 Announce Type: new Abstract: Recent advances in neural rendering have introduced numerous 3D scene representations. Although standard computer vision metrics evaluate the visual qua

model-releasesarxiv-cs-cv
21 Apr 2026
Model Releases

A Note on TurboQuant and the Earlier DRIVE/EDEN Line of Work

DGX agent

arXiv:2604.18555v1 Announce Type: new Abstract: This note clarifies the relationship between the recent TurboQuant work and the earlier DRIVE (NeurIPS 2021) and EDEN (ICML 2022) schemes. DRIVE is a 1-

model-releasesarxiv-cs-lg
21 Apr 2026
Model Releases

A Probabilistic Consensus-Driven Approach for Robust Counterfactual Explanations

DGX agent

arXiv:2604.17494v1 Announce Type: new Abstract: Counterfactual explanations (CFEs) are essential for interpreting black-box models, yet they often become invalid when models are slightly changed. Exis

model-releasesarxiv-cs-lg
21 Apr 2026
Model Releases

A Real-World Grasping-in-Clutter Performance Evaluation Benchmark for Robotic Food Waste Sorting

DGX agent

arXiv:2602.18835v2 Announce Type: replace Abstract: Food waste management is critical for sustainability, yet inorganic contaminants hinder recycling potential. Robotic automation accelerates sorting

model-releasesarxiv-cs-ro
21 Apr 2026
Model Releases

A Survey of Spatial Memory Representations for Efficient Robot Navigation

DGX agent

arXiv:2604.16482v1 Announce Type: new Abstract: As vision-based robots navigate larger environments, their spatial memory grows without bound, eventually exhausting computational resources, particular

model-releasesarxiv-cs-cv
21 Apr 2026
Model Releases

A Systematic Survey and Benchmark of Deep Learning for Molecular Property Prediction in the Foundation Model Era

DGX agent

arXiv:2604.16586v1 Announce Type: new Abstract: Molecular property prediction integrates quantum chemistry, cheminformatics, and deep learning to connect molecular structure with physicochemical and b

model-releasesarxiv-cs-lg
21 Apr 2026
Model Releases

A Transformer and Prototype-based Interpretable Model for Contextual Sarcasm Detection

DGX agent

arXiv:2503.11838v2 Announce Type: replace Abstract: Sarcasm detection, with its figurative nature, poses unique challenges for affective systems designed to perform sentiment analysis. While these sys

model-releasesarxiv-cs-cl
21 Apr 2026
Model Releases

Abstain-R1: Calibrated Abstention and Post-Refusal Clarification via Verifiable RL

DGX agent

arXiv:2604.17073v1 Announce Type: new Abstract: Reinforcement fine-tuning improves the reasoning ability of large language models, but it can also encourage them to answer unanswerable queries by gues

model-releasesarxiv-cs-cl
21 Apr 2026
Model Releases

Adaptive Forensic Feature Refinement via Intrinsic Importance Perception

DGX agent

arXiv:2604.16879v1 Announce Type: new Abstract: With the rapid development of generative models and multimodal content editing technologies, the key challenge faced by synthetic image detection (SID)

model-releasesarxiv-cs-cv
21 Apr 2026
Model Releases

Adaptive Local Frequency Filtering for Fourier-Encoded Implicit Neural Representations

DGX agent

arXiv:2604.02846v2 Announce Type: replace Abstract: Fourier-encoded implicit neural representations (INRs) have shown strong capability in modeling continuous signals from discrete samples. However, c

model-releasesarxiv-cs-cv
21 Apr 2026
Model Releases

Adaptive Text Anonymization: Learning Privacy-Utility Trade-offs via Prompt Optimization

DGX agent

arXiv:2602.20743v2 Announce Type: replace Abstract: Anonymizing textual documents is a highly context-sensitive problem: the appropriate balance between privacy protection and utility preservation var

model-releasesarxiv-cs-cl
21 Apr 2026
Model Releases

Adversarial Humanities Benchmark: Results on Stylistic Robustness in Frontier Model Safety

DGX agent

arXiv:2604.18487v1 Announce Type: new Abstract: The Adversarial Humanities Benchmark (AHB) evaluates whether model safety refusals survive a shift away from familiar harmful prompt forms. Starting fro

model-releasesarxiv-cs-cl
21 Apr 2026
Model Releases

Adverse-to-the-eXtreme Panoptic Segmentation: URVIS 2026 Study and Benchmark

DGX agent

arXiv:2604.16984v1 Announce Type: new Abstract: This paper presents the report of the URVIS 2026 challenge on adverse-to-extreme panoptic segmentation. As the first challenge of its kind, it attracted

model-releasesarxiv-cs-cv
21 Apr 2026
Model Releases

AeroRAG: Structured Multimodal Retrieval-Augmented LLM for Fine-Grained Aerial Visual Reasoning

DGX agent

arXiv:2604.17889v1 Announce Type: new Abstract: Despite recent progress in multimodal large language models (MLLMs), reliable visual question answering in aerial scenes remains challenging. In such sc

model-releasesarxiv-cs-cv
21 Apr 2026
← Previous
1…311312313314315…357
Next →