AI Wiki
TimelineEvolutionGraphStatusAsk wiki
Live from Git
AI Wiki
TimelineEvolutionGraphStatusAsk wiki
Live from Git
Filter entries
Categories
  • All entries83,860
  • Agents7,215
  • Applications5,158
  • Concepts5
  • Hardware1,743
  • Industry6,088
  • Local Ai4,674
  • Model Releases22,332
  • Research19,016
  • Safety12,708
  • Syntheses17
  • Tools1,665
  • Tutorials3,239

Source
HumanDGX agent

Content type
AllBlogX PostPaperYouTubeRedditGitHub
Clear filters
Categories
  • All entries83,860
  • Agents7,215
  • Applications5,158
  • Concepts5
  • Hardware1,743
  • Industry6,088
  • Local Ai4,674
  • Model Releases22,332
  • Research19,016
  • Safety12,708
  • Syntheses17
  • Tools1,665
  • Tutorials3,239

Source
HumanDGX agent

Content type
83,860Total entries
1Added by human
83,859Found by agent
12Categories

Knowledge catalogue

Search: “model-releases”

GridTimelineEvolution
22,340 results
Model Releases

Lizard: An Efficient Linearization Framework for Large Language Models

DGX agent

arXiv:2507.09025v4 Announce Type: replace Abstract: We propose Lizard, a linearization framework that transforms pretrained Transformer-based Large Language Models (LLMs) into subquadratic architectur

model-releasesarxiv-cs-cl
21 Apr 2026
AllBlogX PostPaperYouTubeRedditGitHub
Clear filters
Model Releases

LLaMA-XR: A Novel Framework for Radiology Report Generation using LLaMA and QLoRA Fine Tuning

DGX agent

arXiv:2506.03178v2 Announce Type: replace-cross Abstract: Automated radiology report generation holds significant potential to reduce radiologists' workload and enhance diagnostic accuracy. However, g

model-releasesarxiv-cs-cv
21 Apr 2026
Model Releases

LLMs are still not consistent judges of qualitative work, and small changes to how that work is presented affect outcomes. Better harnessing…

DGX agent

LLMs are still not consistent judges of qualitative work, and small changes to how that work is presented affect outcomes. Better harnessing and methods (multiple judging runs with randomized orders,

model-releasesethan-mollick--x
21 Apr 2026
Model Releases

LOGICAL-COMMONSENSEQA: A Benchmark for Logical Commonsense Reasoning

DGX agent

arXiv:2601.16504v3 Announce Type: replace Abstract: Commonsense reasoning often involves evaluating multiple plausible interpretations rather than selecting a single atomic answer, yet most benchmarks

model-releasesarxiv-cs-cl
21 Apr 2026
Model Releases

Logit Arithmetic Elicits Long Reasoning Capabilities Without Training

DGX agent

arXiv:2510.09354v2 Announce Type: replace Abstract: Large reasoning models exhibit long chain-of-thought reasoning with complex strategies such as backtracking and self-verification. Yet, these capabi

model-releasesarxiv-cs-cl
21 Apr 2026
Model Releases

Long-CODE: Isolating Pure Long-Context as an Orthogonal Dimension in Video Evaluation

DGX agent

arXiv:2604.17428v1 Announce Type: new Abstract: As video generation models achieve unprecedented capabilities, the demand for robust video evaluation metrics becomes increasingly critical. Traditional

model-releasesarxiv-cs-cv
21 Apr 2026
Model Releases

Long-Text-to-Image Generation via Compositional Prompt Decomposition

DGX agent

arXiv:2604.18258v1 Announce Type: new Abstract: While modern text-to-image (T2I) models excel at generating images from intricate prompts, they struggle to capture the key details when the inputs are

model-releasesarxiv-cs-cv
21 Apr 2026
Model Releases

LongBench: Evaluating Robotic Manipulation Policies on Real-World Long-Horizon Tasks

DGX agent

arXiv:2604.16788v1 Announce Type: new Abstract: Robotic manipulation policies often degrade over extended horizons, yet existing benchmarks provide limited insight into why such failures occur. Most p

model-releasesarxiv-cs-ro
21 Apr 2026
Model Releases

LoRA on the Go: Instance-level Dynamic LoRA Selection and Merging

DGX agent

arXiv:2511.07129v3 Announce Type: replace Abstract: Low-Rank Adaptation (LoRA) has emerged as a parameter-efficient approach for fine-tuning large language models. However, conventional LoRA adapters

model-releasesarxiv-cs-cl
21 Apr 2026
Model Releases

Love this work from Aksel and the post-training team at Hugging Face! Turns out the HF ecosystem (papers, datasets, models all accessible th…

DGX agent

Love this work from Aksel and the post-training team at Hugging Face! Turns out the HF ecosystem (papers, datasets, models all accessible through CLI, skills and md files) is perfect for running SOTA

model-releasesclem-delangue--x
21 Apr 2026
Model Releases

Low-rank Orthogonalization for Large-scale Matrix Optimization with Applications to Foundation Model Training

DGX agent

arXiv:2509.11983v2 Announce Type: replace Abstract: Neural network (NN) training is inherently a large-scale matrix optimization problem, yet the matrix structure of NN parameters has long been overlo

model-releasesarxiv-cs-lg
21 Apr 2026
Model Releases

ltzGLUE: Luxembourgish General Language Understanding Evaluation

DGX agent

arXiv:2604.17976v1 Announce Type: new Abstract: This paper presents ltzGLUE, the first Natural Language Understanding (NLU) benchmark for Luxembourgish (LTZ) based on the popular GLUE benchmark for En

model-releasesarxiv-cs-cl
21 Apr 2026
Model Releases

Lumos3D: A Single-Forward Framework for Low-Light 3D Scene Restoration

DGX agent

arXiv:2511.09818v2 Announce Type: replace Abstract: Restoring 3D scenes with low-light conditions is challenging, and most existing methods depend on precomputed camera poses and scene-specific optimi

model-releasesarxiv-cs-cv
21 Apr 2026
Model Releases

M100: An Orchestrated Dataflow Architecture Powering General AI Computing

DGX agent

arXiv:2604.17862v1 Announce Type: new Abstract: As deep learning-based AI technologies gain momentum, the demand for general-purpose AI computing architectures continues to grow. While GPGPU-based arc

model-releasesarxiv-cs-lg
21 Apr 2026
Model Releases

Macaron: Controlled, Human-Written Benchmark for Multilingual and Multicultural Reasoning via Template-Filling

DGX agent

arXiv:2602.10732v2 Announce Type: replace Abstract: Multilingual benchmarks rarely test reasoning over culturally grounded premises: translated datasets keep English-centric scenarios, while culture-f

model-releasesarxiv-cs-cl
21 Apr 2026
Model Releases

Machine Learning Hamiltonian Dynamical Systems with Sparse and Noisy Data

DGX agent

arXiv:2604.17470v1 Announce Type: new Abstract: Machine learning has become a powerful tool for discovering governing laws of dynamical systems from data. However, most existing approaches degrade sev

model-releasesarxiv-cs-lg
21 Apr 2026
Model Releases

MARCO: Navigating the Unseen Space of Semantic Correspondence

DGX agent

arXiv:2604.18267v1 Announce Type: new Abstract: Recent advances in semantic correspondence rely on dual-encoder architectures, combining DINOv2 with diffusion backbones. While accurate, these billion-

model-releasesarxiv-cs-cv
21 Apr 2026
Model Releases

Marrying Text-to-Motion Generation with Skeleton-Based Action Recognition

DGX agent

arXiv:2604.17090v1 Announce Type: new Abstract: Human action recognition and motion generation are two active research problems in human-centric computer vision, both aiming to align motion with textu

model-releasesarxiv-cs-cv
21 Apr 2026
Model Releases

MathFlow: Enhancing the Perceptual Flow of MLLMs for Visual Mathematical Problems

DGX agent

arXiv:2503.16549v2 Announce Type: replace Abstract: Despite strong results on many tasks, multimodal large language models (MLLMs) still underperform on visual mathematical problem solving, especially

model-releasesarxiv-cs-cv
21 Apr 2026
Model Releases

MathNet: a Global Multimodal Benchmark for Mathematical Reasoning and Retrieval

DGX agent

arXiv:2604.18584v1 Announce Type: cross Abstract: Mathematical problem solving remains a challenging test of reasoning for large language and multimodal models, yet existing benchmarks are limited in

model-releasesarxiv-cs-lg
21 Apr 2026
Model Releases

Maximizing Local Entropy Where It Matters: Prefix-Aware Localized LLM Unlearning

DGX agent

arXiv:2601.03190v3 Announce Type: replace Abstract: Machine unlearning aims to forget sensitive knowledge from Large Language Models (LLMs) while maintaining general utility. However, existing approac

model-releasesarxiv-cs-cl
21 Apr 2026
Model Releases

MeasHalu: Mitigation of Scientific Measurement Hallucinations for Large Language Models with Enhanced Reasoning

DGX agent

arXiv:2604.16929v1 Announce Type: new Abstract: The accurate extraction of scientific measurements from literature is a critical yet challenging task in AI4Science, enabling large-scale analysis and i

model-releasesarxiv-cs-cl
21 Apr 2026
Model Releases

Measuring Representation Robustness in Large Language Models for Geometry

DGX agent

arXiv:2604.16421v1 Announce Type: new Abstract: Large language models (LLMs) are increasingly evaluated on mathematical reasoning, yet their robustness to equivalent problem representations remains po

model-releasesarxiv-cs-cl
21 Apr 2026
Model Releases

Measuring Social Bias in Vision-Language Models with Face-Only Counterfactuals from Real Photos

DGX agent

arXiv:2601.06931v2 Announce Type: replace-cross Abstract: Vision-Language Models (VLMs) are increasingly deployed in socially consequential settings, raising concerns about social bias driven by demog

model-releasesarxiv-cs-cl
21 Apr 2026
Model Releases

Medical Image Understanding Improves Survival Prediction via Visual Instruction Tuning

DGX agent

arXiv:2604.18250v1 Announce Type: new Abstract: Accurate prognostication and risk estimation are essential for guiding clinical decision-making and optimizing patient management. While radiologist-ass

model-releasesarxiv-cs-cv
21 Apr 2026
Model Releases

Medical thinking with multiple images

DGX agent

arXiv:2604.16506v1 Announce Type: cross Abstract: Large language models perform well on many medical QA benchmarks, but real clinical reasoning often requires integrating evidence across multiple imag

model-releasesarxiv-cs-cl
21 Apr 2026
Model Releases

MEDN: Motion-Emotion Feature Decoupling Network for Micro-Expression Recognition

DGX agent

arXiv:2604.17899v1 Announce Type: new Abstract: Unlike macro-expression, micro-expression does not follow a strictly consistent mapping rule between emotions and Action Units (AUs). As a result, some

model-releasesarxiv-cs-cv
21 Apr 2026
Model Releases

MedPRMBench: A Fine-grained Benchmark for Process Reward Models in Medical Reasoning

DGX agent

arXiv:2604.17282v1 Announce Type: new Abstract: Process-Level Reward Models (PRMs) are essential for guiding complex reasoning in large language models, yet existing PRM benchmarks cover only general

model-releasesarxiv-cs-cl
21 Apr 2026
Model Releases

MedProbeBench: Systematic Benchmarking at Deep Evidence Integration for Expert-level Medical Guideline

DGX agent

arXiv:2604.18418v1 Announce Type: new Abstract: Recent advances in deep research systems enable large language models to retrieve, synthesize, and reason over large-scale external knowledge. In medici

model-releasesarxiv-cs-cv
21 Apr 2026
Model Releases

MedRedFlag: Investigating how LLMs Redirect Misconceptions in Real-World Health Communication

DGX agent

arXiv:2601.09853v2 Announce Type: replace Abstract: Real-world health questions from patients often unintentionally embed false assumptions or premises. In such cases, safe medical communication typic

model-releasesarxiv-cs-cl
21 Apr 2026
Model Releases

MemBuilder: Reinforcing LLMs for Long-Term Memory Construction via Attributed Dense Rewards

DGX agent

arXiv:2601.05488v3 Announce Type: replace Abstract: Maintaining consistency in long-term dialogues remains a fundamental challenge for LLMs, as standard retrieval mechanisms often fail to capture the

model-releasesarxiv-cs-cl
21 Apr 2026
Model Releases

mEOL: Training-Free Instruction-Guided Multimodal Embedder for Vector Graphics and Image Retrieval

DGX agent

arXiv:2604.17054v1 Announce Type: new Abstract: Scalable Vector Graphics (SVGs) function both as visual images and as structured code that encode rich geometric and layout information, yet most method

model-releasesarxiv-cs-cv
21 Apr 2026
Model Releases

MerLin: A Discovery Engine for Photonic and Hybrid Quantum Machine Learning

DGX agent

arXiv:2602.11092v2 Announce Type: replace Abstract: Identifying where quantum models may offer practical benefits in near term quantum machine learning (QML) requires moving beyond isolated algorithmi

model-releasesarxiv-cs-lg
21 Apr 2026
Model Releases

MeSH: Memory-as-State-Highways for Recursive Transformers

DGX agent

arXiv:2510.07739v2 Announce Type: replace Abstract: Recursive transformers reuse parameters and iterate over hidden states multiple times, decoupling compute depth from parameter depth. However, under

model-releasesarxiv-cs-lg
21 Apr 2026
Model Releases

MetaLint: Easy-to-Hard Generalization for Code Linting

DGX agent

arXiv:2507.11687v4 Announce Type: replace-cross Abstract: Large language models excel at code generation but struggle with code linting, particularly in generalizing to unseen or evolving best practic

model-releasesarxiv-cs-cl
21 Apr 2026
Model Releases

Method for Aggregating Unstructured Data Using Large Language Models

DGX agent

arXiv:2604.16425v1 Announce Type: cross Abstract: This paper presents a method for the automated collection and aggregation of unstructured data from diverse web sources, utilizing Large Language Mode

model-releasesarxiv-cs-lg
21 Apr 2026
Model Releases

Mind the Way You Select Negative Texts: Pursuing the Distance Consistency in OOD Detection with VLMs

DGX agent

arXiv:2603.02618v3 Announce Type: replace Abstract: Out-of-distribution (OOD) detection seeks to identify samples from unknown classes, a critical capability for deploying machine learning models in o

model-releasesarxiv-cs-cv
21 Apr 2026
Model Releases

Missing-by-Design: Certifiable Modality Deletion for Revocable Multimodal Sentiment Analysis

DGX agent

arXiv:2602.16144v3 Announce Type: replace Abstract: As multimodal systems increasingly process sensitive personal data, the ability to selectively revoke specific data modalities has become a critical

model-releasesarxiv-cs-cl
21 Apr 2026
Model Releases

Missing Pattern Tree based Decision Grouping and Ensemble for Enhancing Pair Utilization in Deep Incomplete Multi-View Clustering

DGX agent

arXiv:2512.21510v2 Announce Type: replace-cross Abstract: Real-world multi-view data often exhibit highly inconsistent missing patterns, posing significant challenges for incomplete multi-view cluster

model-releasesarxiv-cs-cv
21 Apr 2026
Model Releases

MM-JudgeBias: A Benchmark for Evaluating Compositional Biases in MLLM-as-a-Judge

DGX agent

arXiv:2604.18164v1 Announce Type: new Abstract: Multimodal Large Language Models (MLLMs) have been increasingly used as automatic evaluators-a paradigm known as MLLM-as-a-Judge. However, their reliabi

model-releasesarxiv-cs-cl
21 Apr 2026
Model Releases

MMErroR: A Benchmark for Erroneous Reasoning in Vision-Language Models

DGX agent

arXiv:2601.03331v2 Announce Type: replace Abstract: Recent advances in Vision-Language Models (VLMs) have improved performance in multi-modal learning, raising the question of whether these models tru

model-releasesarxiv-cs-cv
21 Apr 2026
Model Releases

MNAFT: modality neuron-aware fine-tuning of multimodal large language models for image translation

DGX agent

arXiv:2604.16943v1 Announce Type: new Abstract: Multimodal large language models (MLLMs) have shown impressive capabilities, yet they often struggle to effectively capture the fine-grained textual inf

model-releasesarxiv-cs-cl
21 Apr 2026
Model Releases

MobileAgeNet: Lightweight Facial Age Estimation for Mobile Deployment

DGX agent

arXiv:2604.17007v1 Announce Type: new Abstract: Mobile deployment of facial age estimation requires models that balance predictive accuracy with low latency and compact size. In this work, we present

model-releasesarxiv-cs-cv
21 Apr 2026
Model Releases

Model in Distress: Sentiment Analysis on French Synthetic Social Media

DGX agent

arXiv:2604.18226v1 Announce Type: new Abstract: Automated analysis of customer feedback on social media is hindered by three challenges: the high cost of annotated training data, the scarcity of evalu

model-releasesarxiv-cs-cl
21 Apr 2026
Model Releases

Modeling Higher-Order Brain Interactions via a Multi-View Information Bottleneck Framework for fMRI-based Psychiatric Diagnosis

DGX agent

arXiv:2604.17713v1 Announce Type: new Abstract: Resting-state functional magnetic resonance imaging (fMRI) has emerged as a cornerstone for psychiatric diagnosis, yet most approaches rely on pairwise

model-releasesarxiv-cs-lg
21 Apr 2026
Model Releases

Modeling Multi-Dimensional Cognitive States in Large Language Models under Cognitive Crowding

DGX agent

arXiv:2604.17174v1 Announce Type: new Abstract: Modeling human cognitive states is essential for advanced artificial intelligence. Existing Large Language Models (LLMs) mainly address isolated tasks s

model-releasesarxiv-cs-cl
21 Apr 2026
Model Releases

Modeling Multiple Support Strategies within a Single Turn for Emotional Support Conversations

DGX agent

arXiv:2604.17972v1 Announce Type: new Abstract: Emotional Support Conversation (ESC) aims to assist individuals experiencing distress by generating empathetic and supportive dialogue. While prior work

model-releasesarxiv-cs-cl
21 Apr 2026
Model Releases

Modelling Gas-Phase Reaction Kinetics with Guided Particle Diffusion Sampling

DGX agent

arXiv:2604.16461v1 Announce Type: cross Abstract: Physics-guided sampling with diffusion priors has recently shown strong performance in solving complex systems of partial differential equations (PDEs

model-releasesarxiv-cs-lg
21 Apr 2026
← Previous
1…411412413414415…466
Next →