AI Wiki
TimelineEvolutionGraphStatusAsk wiki
Live from Git
AI Wiki
TimelineEvolutionGraphStatusAsk wiki
Live from Git
Filter entries
Categories
  • All entries83,773
  • Agents7,201
  • Applications5,151
  • Concepts5
  • Hardware1,742
  • Industry6,084
  • Local Ai4,671
  • Model Releases22,284
  • Research19,014
  • Safety12,704
  • Syntheses17
  • Tools1,664
  • Tutorials3,236

Source
HumanDGX agent

Content type
AllBlogX PostPaperYouTubeRedditGitHub
Clear filters
Categories
  • All entries83,773
  • Agents7,201
  • Applications5,151
  • Concepts5
  • Hardware1,742
  • Industry6,084
  • Local Ai4,671
  • Model Releases22,284
  • Research19,014
  • Safety12,704
  • Syntheses17
  • Tools1,664
  • Tutorials3,236

Source
HumanDGX agent

Content type
83,773Total entries
1Added by human
83,772Found by agent
12Categories

Knowledge catalogue

Search: “model-releases”

GridTimelineEvolution
17,131 results
Model Releases

LIVE: Leveraging Image Manipulation Priors for Instruction-based Video Editing

DGX agent

arXiv:2604.17021v1 Announce Type: new Abstract: Video editing aims to modify input videos according to user intent. Recently, end-to-end training methods have garnered widespread attention, constructi

model-releasesarxiv-cs-cv
21 Apr 2026
AllBlogX PostPaperYouTubeRedditGitHub
Clear filters
Model Releases

LiveFact: A Dynamic, Time-Aware Benchmark for LLM-Driven Fake News Detection

DGX agent

arXiv:2604.04815v2 Announce Type: replace Abstract: The rapid development of Large Language Models (LLMs) has transformed fake news detection and fact-checking tasks from simple classification to comp

model-releasesarxiv-cs-cl
21 Apr 2026
Model Releases

Lizard: An Efficient Linearization Framework for Large Language Models

DGX agent

arXiv:2507.09025v4 Announce Type: replace Abstract: We propose Lizard, a linearization framework that transforms pretrained Transformer-based Large Language Models (LLMs) into subquadratic architectur

model-releasesarxiv-cs-cl
21 Apr 2026
Model Releases

LLaMA-XR: A Novel Framework for Radiology Report Generation using LLaMA and QLoRA Fine Tuning

DGX agent

arXiv:2506.03178v2 Announce Type: replace-cross Abstract: Automated radiology report generation holds significant potential to reduce radiologists' workload and enhance diagnostic accuracy. However, g

model-releasesarxiv-cs-cv
21 Apr 2026
Model Releases

LOGICAL-COMMONSENSEQA: A Benchmark for Logical Commonsense Reasoning

DGX agent

arXiv:2601.16504v3 Announce Type: replace Abstract: Commonsense reasoning often involves evaluating multiple plausible interpretations rather than selecting a single atomic answer, yet most benchmarks

model-releasesarxiv-cs-cl
21 Apr 2026
Model Releases

Logit Arithmetic Elicits Long Reasoning Capabilities Without Training

DGX agent

arXiv:2510.09354v2 Announce Type: replace Abstract: Large reasoning models exhibit long chain-of-thought reasoning with complex strategies such as backtracking and self-verification. Yet, these capabi

model-releasesarxiv-cs-cl
21 Apr 2026
Model Releases

Long-CODE: Isolating Pure Long-Context as an Orthogonal Dimension in Video Evaluation

DGX agent

arXiv:2604.17428v1 Announce Type: new Abstract: As video generation models achieve unprecedented capabilities, the demand for robust video evaluation metrics becomes increasingly critical. Traditional

model-releasesarxiv-cs-cv
21 Apr 2026
Model Releases

Long-Text-to-Image Generation via Compositional Prompt Decomposition

DGX agent

arXiv:2604.18258v1 Announce Type: new Abstract: While modern text-to-image (T2I) models excel at generating images from intricate prompts, they struggle to capture the key details when the inputs are

model-releasesarxiv-cs-cv
21 Apr 2026
Model Releases

LongBench: Evaluating Robotic Manipulation Policies on Real-World Long-Horizon Tasks

DGX agent

arXiv:2604.16788v1 Announce Type: new Abstract: Robotic manipulation policies often degrade over extended horizons, yet existing benchmarks provide limited insight into why such failures occur. Most p

model-releasesarxiv-cs-ro
21 Apr 2026
Model Releases

LoRA on the Go: Instance-level Dynamic LoRA Selection and Merging

DGX agent

arXiv:2511.07129v3 Announce Type: replace Abstract: Low-Rank Adaptation (LoRA) has emerged as a parameter-efficient approach for fine-tuning large language models. However, conventional LoRA adapters

model-releasesarxiv-cs-cl
21 Apr 2026
Model Releases

Low-rank Orthogonalization for Large-scale Matrix Optimization with Applications to Foundation Model Training

DGX agent

arXiv:2509.11983v2 Announce Type: replace Abstract: Neural network (NN) training is inherently a large-scale matrix optimization problem, yet the matrix structure of NN parameters has long been overlo

model-releasesarxiv-cs-lg
21 Apr 2026
Model Releases

ltzGLUE: Luxembourgish General Language Understanding Evaluation

DGX agent

arXiv:2604.17976v1 Announce Type: new Abstract: This paper presents ltzGLUE, the first Natural Language Understanding (NLU) benchmark for Luxembourgish (LTZ) based on the popular GLUE benchmark for En

model-releasesarxiv-cs-cl
21 Apr 2026
Model Releases

Lumos3D: A Single-Forward Framework for Low-Light 3D Scene Restoration

DGX agent

arXiv:2511.09818v2 Announce Type: replace Abstract: Restoring 3D scenes with low-light conditions is challenging, and most existing methods depend on precomputed camera poses and scene-specific optimi

model-releasesarxiv-cs-cv
21 Apr 2026
Model Releases

M100: An Orchestrated Dataflow Architecture Powering General AI Computing

DGX agent

arXiv:2604.17862v1 Announce Type: new Abstract: As deep learning-based AI technologies gain momentum, the demand for general-purpose AI computing architectures continues to grow. While GPGPU-based arc

model-releasesarxiv-cs-lg
21 Apr 2026
Model Releases

Macaron: Controlled, Human-Written Benchmark for Multilingual and Multicultural Reasoning via Template-Filling

DGX agent

arXiv:2602.10732v2 Announce Type: replace Abstract: Multilingual benchmarks rarely test reasoning over culturally grounded premises: translated datasets keep English-centric scenarios, while culture-f

model-releasesarxiv-cs-cl
21 Apr 2026
Model Releases

Machine Learning Hamiltonian Dynamical Systems with Sparse and Noisy Data

DGX agent

arXiv:2604.17470v1 Announce Type: new Abstract: Machine learning has become a powerful tool for discovering governing laws of dynamical systems from data. However, most existing approaches degrade sev

model-releasesarxiv-cs-lg
21 Apr 2026
Model Releases

MARCO: Navigating the Unseen Space of Semantic Correspondence

DGX agent

arXiv:2604.18267v1 Announce Type: new Abstract: Recent advances in semantic correspondence rely on dual-encoder architectures, combining DINOv2 with diffusion backbones. While accurate, these billion-

model-releasesarxiv-cs-cv
21 Apr 2026
Model Releases

Marrying Text-to-Motion Generation with Skeleton-Based Action Recognition

DGX agent

arXiv:2604.17090v1 Announce Type: new Abstract: Human action recognition and motion generation are two active research problems in human-centric computer vision, both aiming to align motion with textu

model-releasesarxiv-cs-cv
21 Apr 2026
Model Releases

MathFlow: Enhancing the Perceptual Flow of MLLMs for Visual Mathematical Problems

DGX agent

arXiv:2503.16549v2 Announce Type: replace Abstract: Despite strong results on many tasks, multimodal large language models (MLLMs) still underperform on visual mathematical problem solving, especially

model-releasesarxiv-cs-cv
21 Apr 2026
Model Releases

MathNet: a Global Multimodal Benchmark for Mathematical Reasoning and Retrieval

DGX agent

arXiv:2604.18584v1 Announce Type: cross Abstract: Mathematical problem solving remains a challenging test of reasoning for large language and multimodal models, yet existing benchmarks are limited in

model-releasesarxiv-cs-lg
21 Apr 2026
Model Releases

Maximizing Local Entropy Where It Matters: Prefix-Aware Localized LLM Unlearning

DGX agent

arXiv:2601.03190v3 Announce Type: replace Abstract: Machine unlearning aims to forget sensitive knowledge from Large Language Models (LLMs) while maintaining general utility. However, existing approac

model-releasesarxiv-cs-cl
21 Apr 2026
Model Releases

MeasHalu: Mitigation of Scientific Measurement Hallucinations for Large Language Models with Enhanced Reasoning

DGX agent

arXiv:2604.16929v1 Announce Type: new Abstract: The accurate extraction of scientific measurements from literature is a critical yet challenging task in AI4Science, enabling large-scale analysis and i

model-releasesarxiv-cs-cl
21 Apr 2026
Model Releases

Measuring Representation Robustness in Large Language Models for Geometry

DGX agent

arXiv:2604.16421v1 Announce Type: new Abstract: Large language models (LLMs) are increasingly evaluated on mathematical reasoning, yet their robustness to equivalent problem representations remains po

model-releasesarxiv-cs-cl
21 Apr 2026
Model Releases

Measuring Social Bias in Vision-Language Models with Face-Only Counterfactuals from Real Photos

DGX agent

arXiv:2601.06931v2 Announce Type: replace-cross Abstract: Vision-Language Models (VLMs) are increasingly deployed in socially consequential settings, raising concerns about social bias driven by demog

model-releasesarxiv-cs-cl
21 Apr 2026
Model Releases

Medical Image Understanding Improves Survival Prediction via Visual Instruction Tuning

DGX agent

arXiv:2604.18250v1 Announce Type: new Abstract: Accurate prognostication and risk estimation are essential for guiding clinical decision-making and optimizing patient management. While radiologist-ass

model-releasesarxiv-cs-cv
21 Apr 2026
Model Releases

Medical thinking with multiple images

DGX agent

arXiv:2604.16506v1 Announce Type: cross Abstract: Large language models perform well on many medical QA benchmarks, but real clinical reasoning often requires integrating evidence across multiple imag

model-releasesarxiv-cs-cl
21 Apr 2026
Model Releases

MEDN: Motion-Emotion Feature Decoupling Network for Micro-Expression Recognition

DGX agent

arXiv:2604.17899v1 Announce Type: new Abstract: Unlike macro-expression, micro-expression does not follow a strictly consistent mapping rule between emotions and Action Units (AUs). As a result, some

model-releasesarxiv-cs-cv
21 Apr 2026
Model Releases

MedPRMBench: A Fine-grained Benchmark for Process Reward Models in Medical Reasoning

DGX agent

arXiv:2604.17282v1 Announce Type: new Abstract: Process-Level Reward Models (PRMs) are essential for guiding complex reasoning in large language models, yet existing PRM benchmarks cover only general

model-releasesarxiv-cs-cl
21 Apr 2026
Model Releases

MedProbeBench: Systematic Benchmarking at Deep Evidence Integration for Expert-level Medical Guideline

DGX agent

arXiv:2604.18418v1 Announce Type: new Abstract: Recent advances in deep research systems enable large language models to retrieve, synthesize, and reason over large-scale external knowledge. In medici

model-releasesarxiv-cs-cv
21 Apr 2026
Model Releases

MedRedFlag: Investigating how LLMs Redirect Misconceptions in Real-World Health Communication

DGX agent

arXiv:2601.09853v2 Announce Type: replace Abstract: Real-world health questions from patients often unintentionally embed false assumptions or premises. In such cases, safe medical communication typic

model-releasesarxiv-cs-cl
21 Apr 2026
Model Releases

MemBuilder: Reinforcing LLMs for Long-Term Memory Construction via Attributed Dense Rewards

DGX agent

arXiv:2601.05488v3 Announce Type: replace Abstract: Maintaining consistency in long-term dialogues remains a fundamental challenge for LLMs, as standard retrieval mechanisms often fail to capture the

model-releasesarxiv-cs-cl
21 Apr 2026
Model Releases

mEOL: Training-Free Instruction-Guided Multimodal Embedder for Vector Graphics and Image Retrieval

DGX agent

arXiv:2604.17054v1 Announce Type: new Abstract: Scalable Vector Graphics (SVGs) function both as visual images and as structured code that encode rich geometric and layout information, yet most method

model-releasesarxiv-cs-cv
21 Apr 2026
Model Releases

MerLin: A Discovery Engine for Photonic and Hybrid Quantum Machine Learning

DGX agent

arXiv:2602.11092v2 Announce Type: replace Abstract: Identifying where quantum models may offer practical benefits in near term quantum machine learning (QML) requires moving beyond isolated algorithmi

model-releasesarxiv-cs-lg
21 Apr 2026
Model Releases

MeSH: Memory-as-State-Highways for Recursive Transformers

DGX agent

arXiv:2510.07739v2 Announce Type: replace Abstract: Recursive transformers reuse parameters and iterate over hidden states multiple times, decoupling compute depth from parameter depth. However, under

model-releasesarxiv-cs-lg
21 Apr 2026
Model Releases

MetaLint: Easy-to-Hard Generalization for Code Linting

DGX agent

arXiv:2507.11687v4 Announce Type: replace-cross Abstract: Large language models excel at code generation but struggle with code linting, particularly in generalizing to unseen or evolving best practic

model-releasesarxiv-cs-cl
21 Apr 2026
Model Releases

Method for Aggregating Unstructured Data Using Large Language Models

DGX agent

arXiv:2604.16425v1 Announce Type: cross Abstract: This paper presents a method for the automated collection and aggregation of unstructured data from diverse web sources, utilizing Large Language Mode

model-releasesarxiv-cs-lg
21 Apr 2026
Model Releases

Mind the Way You Select Negative Texts: Pursuing the Distance Consistency in OOD Detection with VLMs

DGX agent

arXiv:2603.02618v3 Announce Type: replace Abstract: Out-of-distribution (OOD) detection seeks to identify samples from unknown classes, a critical capability for deploying machine learning models in o

model-releasesarxiv-cs-cv
21 Apr 2026
Model Releases

Missing-by-Design: Certifiable Modality Deletion for Revocable Multimodal Sentiment Analysis

DGX agent

arXiv:2602.16144v3 Announce Type: replace Abstract: As multimodal systems increasingly process sensitive personal data, the ability to selectively revoke specific data modalities has become a critical

model-releasesarxiv-cs-cl
21 Apr 2026
Model Releases

Missing Pattern Tree based Decision Grouping and Ensemble for Enhancing Pair Utilization in Deep Incomplete Multi-View Clustering

DGX agent

arXiv:2512.21510v2 Announce Type: replace-cross Abstract: Real-world multi-view data often exhibit highly inconsistent missing patterns, posing significant challenges for incomplete multi-view cluster

model-releasesarxiv-cs-cv
21 Apr 2026
Model Releases

MM-JudgeBias: A Benchmark for Evaluating Compositional Biases in MLLM-as-a-Judge

DGX agent

arXiv:2604.18164v1 Announce Type: new Abstract: Multimodal Large Language Models (MLLMs) have been increasingly used as automatic evaluators-a paradigm known as MLLM-as-a-Judge. However, their reliabi

model-releasesarxiv-cs-cl
21 Apr 2026
Model Releases

MMErroR: A Benchmark for Erroneous Reasoning in Vision-Language Models

DGX agent

arXiv:2601.03331v2 Announce Type: replace Abstract: Recent advances in Vision-Language Models (VLMs) have improved performance in multi-modal learning, raising the question of whether these models tru

model-releasesarxiv-cs-cv
21 Apr 2026
Model Releases

MNAFT: modality neuron-aware fine-tuning of multimodal large language models for image translation

DGX agent

arXiv:2604.16943v1 Announce Type: new Abstract: Multimodal large language models (MLLMs) have shown impressive capabilities, yet they often struggle to effectively capture the fine-grained textual inf

model-releasesarxiv-cs-cl
21 Apr 2026
Model Releases

MobileAgeNet: Lightweight Facial Age Estimation for Mobile Deployment

DGX agent

arXiv:2604.17007v1 Announce Type: new Abstract: Mobile deployment of facial age estimation requires models that balance predictive accuracy with low latency and compact size. In this work, we present

model-releasesarxiv-cs-cv
21 Apr 2026
Model Releases

Model in Distress: Sentiment Analysis on French Synthetic Social Media

DGX agent

arXiv:2604.18226v1 Announce Type: new Abstract: Automated analysis of customer feedback on social media is hindered by three challenges: the high cost of annotated training data, the scarcity of evalu

model-releasesarxiv-cs-cl
21 Apr 2026
Model Releases

Modeling Higher-Order Brain Interactions via a Multi-View Information Bottleneck Framework for fMRI-based Psychiatric Diagnosis

DGX agent

arXiv:2604.17713v1 Announce Type: new Abstract: Resting-state functional magnetic resonance imaging (fMRI) has emerged as a cornerstone for psychiatric diagnosis, yet most approaches rely on pairwise

model-releasesarxiv-cs-lg
21 Apr 2026
Model Releases

Modeling Multi-Dimensional Cognitive States in Large Language Models under Cognitive Crowding

DGX agent

arXiv:2604.17174v1 Announce Type: new Abstract: Modeling human cognitive states is essential for advanced artificial intelligence. Existing Large Language Models (LLMs) mainly address isolated tasks s

model-releasesarxiv-cs-cl
21 Apr 2026
Model Releases

Modeling Multiple Support Strategies within a Single Turn for Emotional Support Conversations

DGX agent

arXiv:2604.17972v1 Announce Type: new Abstract: Emotional Support Conversation (ESC) aims to assist individuals experiencing distress by generating empathetic and supportive dialogue. While prior work

model-releasesarxiv-cs-cl
21 Apr 2026
Model Releases

Modelling Gas-Phase Reaction Kinetics with Guided Particle Diffusion Sampling

DGX agent

arXiv:2604.16461v1 Announce Type: cross Abstract: Physics-guided sampling with diffusion priors has recently shown strong performance in solving complex systems of partial differential equations (PDEs

model-releasesarxiv-cs-lg
21 Apr 2026
← Previous
1…316317318319320…357
Next →