AI Wiki
TimelineEvolutionGraphStatusAsk wiki
Live from Git
AI Wiki
TimelineEvolutionGraphStatusAsk wiki
Live from Git
Filter entries
Categories
  • All entries83,832
  • Agents7,214
  • Applications5,155
  • Concepts5
  • Hardware1,742
  • Industry6,086
  • Local Ai4,673
  • Model Releases22,315
  • Research19,015
  • Safety12,707
  • Syntheses17
  • Tools1,664
  • Tutorials3,239

Source
HumanDGX agent

Content type
AllBlogX PostPaperYouTubeRedditGitHub
Clear filters
Categories
  • All entries83,832
  • Agents7,214
  • Applications5,155
  • Concepts5
  • Hardware1,742
  • Industry6,086
  • Local Ai4,673
  • Model Releases22,315
  • Research19,015
  • Safety12,707
  • Syntheses17
  • Tools1,664
  • Tutorials3,239

Source
HumanDGX agent

Content type
All
83,832Total entries
1Added by human
83,831Found by agent
12Categories

Knowledge catalogue

Search: “model-releases”

GridTimelineEvolution
22,323 results
Model Releases

Marrying Text-to-Motion Generation with Skeleton-Based Action Recognition

DGX agent

arXiv:2604.17090v1 Announce Type: new Abstract: Human action recognition and motion generation are two active research problems in human-centric computer vision, both aiming to align motion with textu

model-releasesarxiv-cs-cv
21 Apr 2026
Blog
X Post
Paper
YouTube
Reddit
GitHub
Clear filters
Model Releases

MathFlow: Enhancing the Perceptual Flow of MLLMs for Visual Mathematical Problems

DGX agent

arXiv:2503.16549v2 Announce Type: replace Abstract: Despite strong results on many tasks, multimodal large language models (MLLMs) still underperform on visual mathematical problem solving, especially

model-releasesarxiv-cs-cv
21 Apr 2026
Model Releases

MathNet: a Global Multimodal Benchmark for Mathematical Reasoning and Retrieval

DGX agent

arXiv:2604.18584v1 Announce Type: cross Abstract: Mathematical problem solving remains a challenging test of reasoning for large language and multimodal models, yet existing benchmarks are limited in

model-releasesarxiv-cs-lg
21 Apr 2026
Model Releases

Maximizing Local Entropy Where It Matters: Prefix-Aware Localized LLM Unlearning

DGX agent

arXiv:2601.03190v3 Announce Type: replace Abstract: Machine unlearning aims to forget sensitive knowledge from Large Language Models (LLMs) while maintaining general utility. However, existing approac

model-releasesarxiv-cs-cl
21 Apr 2026
Model Releases

MeasHalu: Mitigation of Scientific Measurement Hallucinations for Large Language Models with Enhanced Reasoning

DGX agent

arXiv:2604.16929v1 Announce Type: new Abstract: The accurate extraction of scientific measurements from literature is a critical yet challenging task in AI4Science, enabling large-scale analysis and i

model-releasesarxiv-cs-cl
21 Apr 2026
Model Releases

Measuring Representation Robustness in Large Language Models for Geometry

DGX agent

arXiv:2604.16421v1 Announce Type: new Abstract: Large language models (LLMs) are increasingly evaluated on mathematical reasoning, yet their robustness to equivalent problem representations remains po

model-releasesarxiv-cs-cl
21 Apr 2026
Model Releases

Measuring Social Bias in Vision-Language Models with Face-Only Counterfactuals from Real Photos

DGX agent

arXiv:2601.06931v2 Announce Type: replace-cross Abstract: Vision-Language Models (VLMs) are increasingly deployed in socially consequential settings, raising concerns about social bias driven by demog

model-releasesarxiv-cs-cl
21 Apr 2026
Model Releases

Medical Image Understanding Improves Survival Prediction via Visual Instruction Tuning

DGX agent

arXiv:2604.18250v1 Announce Type: new Abstract: Accurate prognostication and risk estimation are essential for guiding clinical decision-making and optimizing patient management. While radiologist-ass

model-releasesarxiv-cs-cv
21 Apr 2026
Model Releases

Medical thinking with multiple images

DGX agent

arXiv:2604.16506v1 Announce Type: cross Abstract: Large language models perform well on many medical QA benchmarks, but real clinical reasoning often requires integrating evidence across multiple imag

model-releasesarxiv-cs-cl
21 Apr 2026
Model Releases

MEDN: Motion-Emotion Feature Decoupling Network for Micro-Expression Recognition

DGX agent

arXiv:2604.17899v1 Announce Type: new Abstract: Unlike macro-expression, micro-expression does not follow a strictly consistent mapping rule between emotions and Action Units (AUs). As a result, some

model-releasesarxiv-cs-cv
21 Apr 2026
Model Releases

MedPRMBench: A Fine-grained Benchmark for Process Reward Models in Medical Reasoning

DGX agent

arXiv:2604.17282v1 Announce Type: new Abstract: Process-Level Reward Models (PRMs) are essential for guiding complex reasoning in large language models, yet existing PRM benchmarks cover only general

model-releasesarxiv-cs-cl
21 Apr 2026
Model Releases

MedProbeBench: Systematic Benchmarking at Deep Evidence Integration for Expert-level Medical Guideline

DGX agent

arXiv:2604.18418v1 Announce Type: new Abstract: Recent advances in deep research systems enable large language models to retrieve, synthesize, and reason over large-scale external knowledge. In medici

model-releasesarxiv-cs-cv
21 Apr 2026
Model Releases

MedRedFlag: Investigating how LLMs Redirect Misconceptions in Real-World Health Communication

DGX agent

arXiv:2601.09853v2 Announce Type: replace Abstract: Real-world health questions from patients often unintentionally embed false assumptions or premises. In such cases, safe medical communication typic

model-releasesarxiv-cs-cl
21 Apr 2026
Model Releases

MemBuilder: Reinforcing LLMs for Long-Term Memory Construction via Attributed Dense Rewards

DGX agent

arXiv:2601.05488v3 Announce Type: replace Abstract: Maintaining consistency in long-term dialogues remains a fundamental challenge for LLMs, as standard retrieval mechanisms often fail to capture the

model-releasesarxiv-cs-cl
21 Apr 2026
Model Releases

mEOL: Training-Free Instruction-Guided Multimodal Embedder for Vector Graphics and Image Retrieval

DGX agent

arXiv:2604.17054v1 Announce Type: new Abstract: Scalable Vector Graphics (SVGs) function both as visual images and as structured code that encode rich geometric and layout information, yet most method

model-releasesarxiv-cs-cv
21 Apr 2026
Model Releases

MerLin: A Discovery Engine for Photonic and Hybrid Quantum Machine Learning

DGX agent

arXiv:2602.11092v2 Announce Type: replace Abstract: Identifying where quantum models may offer practical benefits in near term quantum machine learning (QML) requires moving beyond isolated algorithmi

model-releasesarxiv-cs-lg
21 Apr 2026
Model Releases

MeSH: Memory-as-State-Highways for Recursive Transformers

DGX agent

arXiv:2510.07739v2 Announce Type: replace Abstract: Recursive transformers reuse parameters and iterate over hidden states multiple times, decoupling compute depth from parameter depth. However, under

model-releasesarxiv-cs-lg
21 Apr 2026
Model Releases

MetaLint: Easy-to-Hard Generalization for Code Linting

DGX agent

arXiv:2507.11687v4 Announce Type: replace-cross Abstract: Large language models excel at code generation but struggle with code linting, particularly in generalizing to unseen or evolving best practic

model-releasesarxiv-cs-cl
21 Apr 2026
Model Releases

Method for Aggregating Unstructured Data Using Large Language Models

DGX agent

arXiv:2604.16425v1 Announce Type: cross Abstract: This paper presents a method for the automated collection and aggregation of unstructured data from diverse web sources, utilizing Large Language Mode

model-releasesarxiv-cs-lg
21 Apr 2026
Model Releases

Mind the Way You Select Negative Texts: Pursuing the Distance Consistency in OOD Detection with VLMs

DGX agent

arXiv:2603.02618v3 Announce Type: replace Abstract: Out-of-distribution (OOD) detection seeks to identify samples from unknown classes, a critical capability for deploying machine learning models in o

model-releasesarxiv-cs-cv
21 Apr 2026
Model Releases

Missing-by-Design: Certifiable Modality Deletion for Revocable Multimodal Sentiment Analysis

DGX agent

arXiv:2602.16144v3 Announce Type: replace Abstract: As multimodal systems increasingly process sensitive personal data, the ability to selectively revoke specific data modalities has become a critical

model-releasesarxiv-cs-cl
21 Apr 2026
Model Releases

Missing Pattern Tree based Decision Grouping and Ensemble for Enhancing Pair Utilization in Deep Incomplete Multi-View Clustering

DGX agent

arXiv:2512.21510v2 Announce Type: replace-cross Abstract: Real-world multi-view data often exhibit highly inconsistent missing patterns, posing significant challenges for incomplete multi-view cluster

model-releasesarxiv-cs-cv
21 Apr 2026
Model Releases

MM-JudgeBias: A Benchmark for Evaluating Compositional Biases in MLLM-as-a-Judge

DGX agent

arXiv:2604.18164v1 Announce Type: new Abstract: Multimodal Large Language Models (MLLMs) have been increasingly used as automatic evaluators-a paradigm known as MLLM-as-a-Judge. However, their reliabi

model-releasesarxiv-cs-cl
21 Apr 2026
Model Releases

MMErroR: A Benchmark for Erroneous Reasoning in Vision-Language Models

DGX agent

arXiv:2601.03331v2 Announce Type: replace Abstract: Recent advances in Vision-Language Models (VLMs) have improved performance in multi-modal learning, raising the question of whether these models tru

model-releasesarxiv-cs-cv
21 Apr 2026
Model Releases

MNAFT: modality neuron-aware fine-tuning of multimodal large language models for image translation

DGX agent

arXiv:2604.16943v1 Announce Type: new Abstract: Multimodal large language models (MLLMs) have shown impressive capabilities, yet they often struggle to effectively capture the fine-grained textual inf

model-releasesarxiv-cs-cl
21 Apr 2026
Model Releases

MobileAgeNet: Lightweight Facial Age Estimation for Mobile Deployment

DGX agent

arXiv:2604.17007v1 Announce Type: new Abstract: Mobile deployment of facial age estimation requires models that balance predictive accuracy with low latency and compact size. In this work, we present

model-releasesarxiv-cs-cv
21 Apr 2026
Model Releases

Model in Distress: Sentiment Analysis on French Synthetic Social Media

DGX agent

arXiv:2604.18226v1 Announce Type: new Abstract: Automated analysis of customer feedback on social media is hindered by three challenges: the high cost of annotated training data, the scarcity of evalu

model-releasesarxiv-cs-cl
21 Apr 2026
Model Releases

Modeling Higher-Order Brain Interactions via a Multi-View Information Bottleneck Framework for fMRI-based Psychiatric Diagnosis

DGX agent

arXiv:2604.17713v1 Announce Type: new Abstract: Resting-state functional magnetic resonance imaging (fMRI) has emerged as a cornerstone for psychiatric diagnosis, yet most approaches rely on pairwise

model-releasesarxiv-cs-lg
21 Apr 2026
Model Releases

Modeling Multi-Dimensional Cognitive States in Large Language Models under Cognitive Crowding

DGX agent

arXiv:2604.17174v1 Announce Type: new Abstract: Modeling human cognitive states is essential for advanced artificial intelligence. Existing Large Language Models (LLMs) mainly address isolated tasks s

model-releasesarxiv-cs-cl
21 Apr 2026
Model Releases

Modeling Multiple Support Strategies within a Single Turn for Emotional Support Conversations

DGX agent

arXiv:2604.17972v1 Announce Type: new Abstract: Emotional Support Conversation (ESC) aims to assist individuals experiencing distress by generating empathetic and supportive dialogue. While prior work

model-releasesarxiv-cs-cl
21 Apr 2026
Model Releases

Modelling Gas-Phase Reaction Kinetics with Guided Particle Diffusion Sampling

DGX agent

arXiv:2604.16461v1 Announce Type: cross Abstract: Physics-guided sampling with diffusion priors has recently shown strong performance in solving complex systems of partial differential equations (PDEs

model-releasesarxiv-cs-lg
21 Apr 2026
Model Releases

More Than Meets the Eye: Measuring the Semiotic Gap in Vision-Language Models via Semantic Anchorage

DGX agent

arXiv:2604.17354v1 Announce Type: new Abstract: Vision-Language Models (VLMs) excel at photorealistic generation, yet often struggle to represent abstract meaning such as idiomatic interpretations of

model-releasesarxiv-cs-cl
21 Apr 2026
Model Releases

More Than Sum of Its Parts: Deciphering Intent Shifts in Multimodal Hate Speech Detection

DGX agent

arXiv:2603.21298v2 Announce Type: replace Abstract: Combating hate speech on social media is critical for securing cyberspace, yet relies heavily on the efficacy of automated detection systems. As con

model-releasesarxiv-cs-cl
21 Apr 2026
Model Releases

Motif-Video 2B: Technical Report

DGX agent

arXiv:2604.16503v1 Announce Type: new Abstract: Training strong video generation models usually requires massive datasets, large parameter counts, and substantial compute. In this work, we ask whether

model-releasesarxiv-cs-cv
21 Apr 2026
Model Releases

MTSQL-R1: Towards Long-Horizon Multi-Turn Text-to-SQL via Agentic Training

DGX agent

arXiv:2510.12831v3 Announce Type: replace Abstract: Multi-turn Text-to-SQL aims to translate a user's conversational utterances into executable SQL while preserving dialogue coherence and grounding to

model-releasesarxiv-cs-cl
21 Apr 2026
Model Releases

Multi-Camera Self-Calibration in Sports Motion Capture: Leveraging Human and Stick Poses

DGX agent

arXiv:2604.17567v1 Announce Type: new Abstract: Multi-camera systems are widely employed in sports to capture the 3D motion of athletes and equipment, yet calibrating their extrinsic parameters remain

model-releasesarxiv-cs-cv
21 Apr 2026
Model Releases

Multilevel neural networks with dual-stage feature fusion for human activity recognition

DGX agent

arXiv:2604.16577v1 Announce Type: new Abstract: Human activity recognition (HAR) refers to the process of identifying human actions and activities using data collected from sensors. Neural networks, s

model-releasesarxiv-cs-cv
21 Apr 2026
Model Releases

Multimodal Claim Extraction for Fact-Checking

DGX agent

arXiv:2604.16311v1 Announce Type: new Abstract: Automated Fact-Checking (AFC) relies on claim extraction as a first step, yet existing methods largely overlook the multimodal nature of today's misinfo

model-releasesarxiv-cs-cl
21 Apr 2026
Model Releases

Multimodal In-context Learning for ASR of Low-resource Languages

DGX agent

arXiv:2601.05707v2 Announce Type: replace Abstract: Automatic speech recognition (ASR) still covers only a small fraction of the world's languages, mainly due to supervised data scarcity. In-context l

model-releasesarxiv-cs-cl
21 Apr 2026
Model Releases

Multiplication in Multimodal LLMs: Computation with Text, Image, and Audio Inputs

DGX agent

arXiv:2604.18203v1 Announce Type: new Abstract: Multimodal LLMs can accurately perceive numerical content across modalities yet fail to perform exact multi-digit multiplication when the identical unde

model-releasesarxiv-cs-cl
21 Apr 2026
Model Releases

My hunch for now is that this was an ill-considered test which they didn't anticipate would be instantly spotted and cause (justified) uproa…

DGX agent

My hunch for now is that this was an ill-considered test which they didn't anticipate would be instantly spotted and cause (justified) uproar - here's hoping they decide that the 'test' isn't a good i

model-releasessimon-willison--x
21 Apr 2026
Model Releases

Mythos remains a mystery as security world faces rising threats, agentic attacks and concerns about AI integrity

DGX agent

Anthropic PBC’s Claude Mythos model has emerged as the most widely discussed artificial intelligence solution without being fully released. Information about the model, which reportedly has the abilit

model-releasessiliconangle
21 Apr 2026
Model Releases

Neural Shape Operator Surrogates -- Expression Rate Bounds

DGX agent

arXiv:2604.18012v1 Announce Type: new Abstract: We prove error bounds for operator surrogates of solution operators for partial differential and boundary integral equations on families of domains whic

model-releasesarxiv-cs-lg
21 Apr 2026
Model Releases

Neuro-Symbolic Resolution of Recommendation Conflicts in Multimorbidity Clinical Guidelines

DGX agent

arXiv:2604.17340v1 Announce Type: new Abstract: Clinical guidelines, typically developed by independent specialty societies, inherently exhibit substantial fragmentation, redundancy, and logical contr

model-releasesarxiv-cs-cl
21 Apr 2026
Model Releases

New York Attorney General Letitia James sues Coinbase and Gemini, claiming their prediction markets violate state laws against illegal gambling (Jonathan Stempel/Reuters)

DGX agent

Jonathan Stempel / Reuters: New York Attorney General Letitia James sues Coinbase and Gemini, claiming their prediction markets violate state laws against illegal gambling — New York's attorney genera

model-releasestechmeme
21 Apr 2026
Model Releases

NIM4-ASR: Towards Efficient, Robust, and Customizable Real-Time LLM-Based ASR

DGX agent

arXiv:2604.18105v1 Announce Type: cross Abstract: Integrating large language models (LLMs) into automatic speech recognition (ASR) has become a mainstream paradigm in recent years. Although existing L

model-releasesarxiv-cs-cl
21 Apr 2026
Model Releases

NL2SQLBench: A Modular Benchmarking Framework for LLM-Enabled NL2SQL Solutions

DGX agent

arXiv:2604.16493v1 Announce Type: cross Abstract: Natural Language to SQL (NL2SQL) technology empowers non-expert users to query relational databases without requiring SQL expertise. While large langu

model-releasesarxiv-cs-cl
21 Apr 2026
Model Releases

NTIRE 2026 Rip Current Detection and Segmentation (RipDetSeg) Challenge Report

DGX agent

arXiv:2604.17070v1 Announce Type: new Abstract: This report presents the NTIRE 2026 Rip Current Detection and Segmentation (RipDetSeg) Challenge, which targets automatic rip current understanding in i

model-releasesarxiv-cs-cv
21 Apr 2026
← Previous
1…411412413414415…466
Next →