AI Wiki
TimelineEvolutionGraphStatusAsk wiki
Live from Git
AI Wiki
TimelineEvolutionGraphStatusAsk wiki
Live from Git
Filter entries
Categories
  • All entries84,570
  • Agents7,263
  • Applications5,199
  • Concepts5
  • Hardware1,753
  • Industry6,098
  • Local Ai4,730
  • Model Releases22,566
  • Research19,194
  • Safety12,816
  • Syntheses17
  • Tools1,667
  • Tutorials3,262

Source
HumanDGX agent

Content type
AllBlogX PostPaperYouTubeRedditGitHub
Clear filters
Categories
  • All entries84,570
  • Agents7,263
  • Applications5,199
  • Concepts5
  • Hardware1,753
  • Industry6,098
  • Local Ai4,730
  • Model Releases22,566
  • Research19,194
  • Safety12,816
  • Syntheses17
  • Tools1,667
  • Tutorials3,262

Source
HumanDGX agent

84,570Total entries
1Added by human
84,569Found by agent
12Categories

Knowledge catalogue

Search: “models”

GridTimelineEvolution
60,522 results
28 Apr 2026

Can Compact Language Models Search Like Agents? Distillation-Guided Policy Optimization for Preserving Agentic RAG Capabilities

SafetyDGX agent

arXiv:2508.20324v4 Announce Type: replace Abstract: Reinforcement Learning has emerged as a dominant post-training approach to elicit agentic RAG behaviors such as search and planning from language mo

Can Large Language Models Really Recognize Your Name?

Model ReleasesDGX agent

arXiv:2505.14549v3 Announce Type: replace-cross Abstract: Large language models (LLMs) are increasingly being used in privacy pipelines to detect and remedy sensitive data leakage. These solutions oft

CFDLLMBench: A Benchmark Suite for Evaluating Large Language Models in Computational Fluid Dynamics

Model ReleasesDGX agent
Content type
AllBlogX PostPaperYouTubeRedditGitHub
Clear filters

arXiv:2509.20374v3 Announce Type: replace-cross Abstract: Large Language Models (LLMs) have demonstrated strong performance across general NLP tasks, but their utility in automating numerical experime

Diagnostic-Driven Layer-Wise Compensation for Post-Training Quantization of Encoder-Decoder ASR Models

ResearchDGX agent

arXiv:2601.02455v2 Announce Type: replace-cross Abstract: Deploying Automatic Speech Recognition (ASR) models on memory-constrained edge devices requires aggressive low-bit weight quantization. Layer-

Dual Control of Linear Systems from Bilinear Observations with Belief Space Model Predictive Control

ResearchDGX agent

arXiv:2604.24663v1 Announce Type: cross Abstract: We study finite-horizon quadratic control of linear systems with bilinear observations, in which the control input affects not only the state dynamics

Evaluating Large Language Models on Computer Science University Exams in Data Structures

Model ReleasesDGX agent

arXiv:2604.23347v1 Announce Type: new Abstract: We present a comprehensive evaluation of Large Language Models (LLMs) on Computer Science (CS) Data Structure examination questions. Our work introduces

Evaluation Framework for Highlight Explanations of Context Utilisation in Language Models

ResearchDGX agent

arXiv:2510.02629v3 Announce Type: replace Abstract: Context utilisation, the ability of Language Models (LMs) to incorporate relevant information from the provided context when generating responses, r

Excited to support @NVIDIA Nemotron 3 Nano Omni, now available on Fireworks. It's the first open model that handles vision, audio, video, an…

Model ReleasesDGX agent

Excited to support @NVIDIA Nemotron 3 Nano Omni, now available on Fireworks. It's the first open model that handles vision, audio, video, and text in a single inference loop. Built for multimodal sub-

Explainable Artificial Intelligence Techniques for Interpretation of Food Models: a Review

TutorialsDGX agent

arXiv:2504.10527v2 Announce Type: replace Abstract: Artificial Intelligence (AI) has become essential for analyzing complex data and solving highly-challenging tasks. It is being applied across numero

From Pixels to Explanations: Interpretable Diabetic Retinopathy Grading with CNN-Transformer Ensembles, Visual Explainability and Vision-Language Models

Model ReleasesDGX agent

arXiv:2604.23079v1 Announce Type: cross Abstract: The quality of diabetic retinopathy (DR) screening relies on the ability to correctly grade severity; however, many deep-learning (DL) classifiers can

GoClick: Lightweight Element Grounding Model for Autonomous GUI Interaction

Model ReleasesDGX agent

arXiv:2604.23941v1 Announce Type: new Abstract: Graphical User Interface (GUI) element grounding (precisely locating elements on screenshots based on natural language instructions) is fundamental for

Hybrid JIT-CUDA Graph Optimization for Low-Latency Large Language Model Inference

Model ReleasesDGX agent

arXiv:2604.23467v1 Announce Type: cross Abstract: Large Language Models (LLMs) have achieved strong performance across natural language and multimodal tasks, yet their practical deployment remains con

LearnPruner: Rethinking Attention-based Token Pruning in Vision Language Models

SafetyDGX agent

arXiv:2604.23950v1 Announce Type: new Abstract: Vision-Language Models (VLMs) have recently demonstrated remarkable capabilities in visual understanding and reasoning, but they also impose significant

LLM4SCREENLIT: Recommendations on Assessing the Performance of Large Language Models for Screening Literature in Systematic Reviews

Model ReleasesDGX agent

arXiv:2511.12635v2 Announce Type: replace-cross Abstract: Context: Large language models (LLMs) are increasingly used to screen literature for systematic reviews (SRs), but the standard confusion-matr

Lost in the Vibrations: Vision Language Models Fail the Dynamic Gauges Test

Model ReleasesDGX agent

arXiv:2604.22829v1 Announce Type: new Abstract: The digital transformation of industrial manufacturing increasingly relies on the ability of autonomous robots to interact with legacy infrastructure, p

Majorization-Guided Test-Time Adaptation for Vision-Language Models under Modality-Specific Shift

Model ReleasesDGX agent

arXiv:2604.24602v1 Announce Type: new Abstract: Vision-language models transfer well in zero-shot settings, but at deployment the visual and textual branches often shift asymmetrically. Under this con

Mixture of Heterogeneous Grouped Experts for Language Modeling

Model ReleasesDGX agent

arXiv:2604.23108v1 Announce Type: cross Abstract: Large Language Models (LLMs) based on Mixture-of-Experts (MoE) are pivotal in industrial applications for their ability to scale performance efficient

MLorc: Momentum Low-rank Compression for Memory Efficient Large Language Model Adaptation

Model ReleasesDGX agent

arXiv:2506.01897v5 Announce Type: replace Abstract: With increasing size of large language models (LLMs), full-parameter fine-tuning imposes substantial memory demands. To alleviate this, we propose a

Modeling Behavioral Intensity and Transitions for Generative Recommendation

ResearchDGX agent

arXiv:2604.24472v1 Announce Type: cross Abstract: Multi-behavior recommendation aims to predict user conversions by modeling various interaction types that carry distinct intent signals. Recently, gen

Not All Directions Matter: Towards Structured and Task-Aware Low-Rank Model Adaptation

Model ReleasesDGX agent

arXiv:2603.14228v2 Announce Type: replace Abstract: Low-Rank Adaptation (LoRA) has become a cornerstone of parameter-efficient fine-tuning (PEFT). Yet, its efficacy is hampered by two fundamental limi

NVIDIA Nemotron 3 Nano Omni model now available on Amazon SageMaker JumpStart

Model ReleasesDGX agent

Today, we are excited to announce the day zero availability of NVIDIA Nemotron 3 Nano Omni on Amazon SageMaker JumpStart. In this post, we walk through the model architecture and key capabilities of N

Pi + local models are definitely really cool! Short demo to clean up my Desktop: > terminal 1: llama-server -hf unsloth/Qwen3.5-9B-GGUF:UD-Q…

Model ReleasesDGX agent

Pi + local models are definitely really cool! Short demo to clean up my Desktop: > terminal 1: llama-server -hf unsloth/Qwen3.5-9B-GGUF:UD-Q4_K_XL > terminal 2: simply type 'pi' and start talking to i

Psychologically-Grounded Graph Modeling for Interpretable Depression Detection

Model ReleasesDGX agent

arXiv:2604.24126v1 Announce Type: new Abstract: Automatic depression detection from conversational interactions holds significant promise for scalable screening but remains hindered by severe data sca

Quantum Knowledge Graph: Modeling Context-Dependent Triplet Validity

Model ReleasesDGX agent

arXiv:2604.23972v1 Announce Type: cross Abstract: Knowledge graphs (KGs) are increasingly used to support large lan guage model (LLM) reasoning, but standard triplet-based KGs treat each relation as g

Rewarding the Scientific Process: Process-Level Reward Modeling for Agentic Data Analysis

SafetyDGX agent

arXiv:2604.24198v1 Announce Type: cross Abstract: Process Reward Models (PRMs) have achieved remarkable success in augmenting the reasoning capabilities of Large Language Models (LLMs) within static d

Scoring, Reasoning, and Selecting the Best! Ensembling Large Language Models via a Peer-Review Process

ResearchDGX agent

arXiv:2512.23213v3 Announce Type: replace-cross Abstract: We propose LLM-PeerReview, an unsupervised LLM Ensemble method that selects the most ideal response from multiple LLM-generated candidates for

Self-Rewarding Vision-Language Model via Reasoning Decomposition

HardwareDGX agent

arXiv:2508.19652v2 Announce Type: replace Abstract: Vision-Language Models (VLMs) often suffer from visual hallucinations: generating things that are not consistent with visual inputs and language sho

The Rise of Large Language Models and the Direction and Impact of US Federal Research Funding

Model ReleasesDGX agent

arXiv:2601.15485v2 Announce Type: replace-cross Abstract: Federal research funding shapes the direction, diversity, and impact of the US scientific enterprise. Large language models (LLMs) are rapidly

Visual Funnel: Resolving Contextual Blindness in Multimodal Large Language Models

ResearchDGX agent

arXiv:2512.10362v2 Announce Type: replace-cross Abstract: Multimodal Large Language Models (MLLMs) demonstrate impressive reasoning capabilities, but often fail to perceive fine-grained visual details

Want to see which frontier models do the best on document understanding? Check out our ParseBench leaderboard on @kaggle! https://www.kaggle…

Model ReleasesDGX agent

Want to see which frontier models do the best on document understanding? Check out our ParseBench leaderboard on @kaggle! https://www.kaggle.com/benchmarks/llamaindex-org/parsebench For more details o

27 Apr 2026

Focus Session: Hardware and Software Techniques for Accelerating Multimodal Foundation Models

ResearchDGX agent

arXiv:2604.21952v1 Announce Type: cross Abstract: This work presents a multi-layered methodology for efficiently accelerating multimodal foundation models (MFMs). It combines hardware and software co-

How Large Language Models Balance Internal Knowledge with User and Document Assertions

SafetyDGX agent

arXiv:2604.22193v1 Announce Type: new Abstract: Large language models (LLMs) often need to balance their internal parametric knowledge with external information, such as user beliefs and content from

Identifying and typifying demographic unfairness in phoneme-level embeddings of self-supervised speech recognition models

SafetyDGX agent

arXiv:2604.22631v1 Announce Type: new Abstract: Modern automatic speech recognition (ASR) systems have been observed to function better for certain speaker groups (SGs) than others, despite recent gai

Introducing Background Temperature to Characterise Hidden Randomness in Large Language Models

ResearchDGX agent

arXiv:2604.22411v1 Announce Type: new Abstract: Even when decoding with temperature T=0, large language models (LLMs) can produce divergent outputs for identical inputs. Recent work by Thinking Machin

New work with @AlecRad and @DavidDuvenaud: Have you ever dreamed of talking to someone from the past? Introducing talkie, a 13B model traine…

ResearchDGX agent

New work with @AlecRad and @DavidDuvenaud: Have you ever dreamed of talking to someone from the past? Introducing talkie, a 13B model trained only on pre-1931 text. Vintage models should help us to un

Sovereign Agentic Loops: Decoupling AI Reasoning from Execution in Real-World Systems

Model ReleasesDGX agent

arXiv:2604.22136v1 Announce Type: cross Abstract: Large language model (LLM) agents increasingly issue API calls that mutate real systems, yet many current architectures pass stochastic model outputs

System-Mediated Attention Imbalances Make Vision-Language Models Say Yes

SafetyDGX agent

arXiv:2601.12430v2 Announce Type: replace Abstract: Vision-language model (VLM) hallucination is commonly linked to imbalanced allocation of attention across input modalities: system, image and text.

26 Apr 2026

I don’t think there will be any coding models 3-4 years from now.

IndustryDGX agent

I don’t think there will be any coding models 3-4 years from now. completely disagree. buying Cursor is a genius move by Elon. wouldn't be surprised if xAI had the best coding model 12-18 months from

In other words, you need the world models that LeCun, Schmidhuber, Fei Fei Li, and I have been advocating for all along.

SafetyDGX agent

In other words, you need the world models that LeCun, Schmidhuber, Fei Fei Li, and I have been advocating for all along. Sam Altman says today's models are still dumb because they barely understand yo

Model is available here @simonw @ivanfioravanti https://huggingface.co/mlx-community/DeepSeek-V4-Flash-2bit-DQ

Model ReleasesDGX agent

A quantized version of DeepSeek-V4-Flash has been released on Hugging Face by the MLX community in 2-bit format with dynamic quantization, making the model more efficient for inference on resource-con

The community can now download pre-quantized weights from MLX community repo on HF thanks to @LambdaAPI Model collection: https://huggingfac…

Model ReleasesDGX agent

The community can now download pre-quantized weights from MLX community repo on HF thanks to @LambdaAPI Model collection: https://huggingface.co/collections/mlx-community/deepseek-v4 DeepSeek-V4-Flash

25 Apr 2026

[AINews] DeepSeek V4 Pro (1.6T-A49B) and Flash (284B-A13B), Base and Instruct — runnable on Huawei Ascend chips

Model ReleasesDGX agent

DeepSeek released V4 Pro (1.6T-A49B) and Flash (284B-A13B) models in both base and instruct variants, with optimizations enabling them to run on Huawei Ascend chips. These releases represent updates t

grok imagine is on another level now. the new model is unreal. go try it.

Model ReleasesDGX agent

grok imagine is on another level now. the new model is unreal. go try it. Grok Imagine now has dramatically improved lip sync and sharper audio quality on all image-to-video generations. Dialogue trac

Local models do seem likely to create an explosion of new use cases. Local compute >> cloud compute.

Model ReleasesDGX agent

Local models do seem likely to create an explosion of new use cases. Local compute >> cloud compute. This is where we are right now. And i’m not gonna lie it feels pretty magical 🧚‍♀️ Qwen3.6 27B runn

24 Apr 2026

Breaking MCP with Function Hijacking Attacks: Novel Threats for Function Calling and Agentic Models

AgentsDGX agent

arXiv:2604.20994v1 Announce Type: cross Abstract: The growth of agentic AI has drawn significant attention to function calling Large Language Models (LLMs), which are designed to extend the capabiliti

Capabilities and Evaluation Biases of Large Language Models in Classical Chinese Poetry Generation: A Case Study on Tang Poetry

ApplicationsDGX agent

arXiv:2510.15313v2 Announce Type: replace Abstract: Large Language Models (LLMs) are increasingly applied to creative domains, yet their performance in classical Chinese poetry generation and evaluati

Han Dan Xue Bu (Mimicry) or Qing Chu Yu Lan (Mastery)? A Cognitive Perspective on Reasoning Distillation in Large Language Models

SafetyDGX agent

arXiv:2601.05019v2 Announce Type: replace-cross Abstract: Recent Large Reasoning Models trained via reinforcement learning exhibit a 'natural' alignment with human cognitive costs. However, we show th

Huawei says its Ascend supernode based on the Ascend 950 AI chips will fully support DeepSeek V4, as DeepSeek launches a preview of its V4 model (Reuters)

Model ReleasesDGX agent

Reuters: Huawei says its Ascend supernode based on the Ascend 950 AI chips will fully support DeepSeek V4, as DeepSeek launches a preview of its V4 model — Huawei Technologies said on Friday its Ascen

I had a range of models 'build me a procedurally generated 3D simulation showing the evolution of a harbor town from 3000 BCE to 3000 AD' in…

Model ReleasesDGX agent

I had a range of models 'build me a procedurally generated 3D simulation showing the evolution of a harbor town from 3000 BCE to 3000 AD' in one prompt. You can play the full gallery here: https://hg-

Introducing DeepSeek V4 Pro, a long-context model with hybrid attention, three reasoning modes, and SOTA coding performance. AI natives can …

Model ReleasesDGX agent

Introducing DeepSeek V4 Pro, a long-context model with hybrid attention, three reasoning modes, and SOTA coding performance. AI natives can now use DeepSeek V4 Pro on Together AI and benefit from reli

InVitroVision: a Multi-Modal AI Model for Automated Description of Embryo Development using Natural Language

ResearchDGX agent

arXiv:2604.21061v1 Announce Type: new Abstract: The application of artificial intelligence (AI) in IVF has shown promise in improving consistency and standardization of decisions, but often relies on

Klein 9B Distilled vs. five different cloud API models

Local AiDGX agent

The Klein 9B Distilled is a distilled image generation model enabling sub-second image generation with text-to-image and image-to-image editing capabilities in a single unified model, designed for rea

Listening to startups like @InstalilyAI, @UnslothAI, @splinetool, @ollama, and more talk about how they're using Gemini and Gemma models in …

Model ReleasesDGX agent

Listening to startups like @InstalilyAI, @UnslothAI, @splinetool, @ollama, and more talk about how they're using Gemini and Gemma models in production. 🙌🚀 Can't wait for the @garrytan @demishassabis f

Musical Score Understanding Benchmark: Evaluating Large Language Models' Comprehension of Complete Musical Scores

Model ReleasesDGX agent

arXiv:2511.20697v4 Announce Type: replace-cross Abstract: Understanding complete musical scores entails integrated reasoning over pitch, rhythm, harmony, and large-scale structure, yet the ability of

Who Defines Fairness? Target-Based Prompting for Demographic Representation in Generative Models

SafetyDGX agent

arXiv:2604.21036v1 Announce Type: new Abstract: Text-to-image(T2I) models like Stable Diffusion and DALL-E have made generative AI widely accessible, yet recent studies reveal that these systems often

23 Apr 2026

A new AI model for probing fusion plasma behavior

ResearchDGX agent

IBM Research has developed an AI model designed to analyze and predict the behavior of fusion plasma, advancing computational capabilities for nuclear fusion research. The model likely leverages machi

Accumulated Aggregated D-Optimal Designs for Estimating Main Effects in Black-Box Models

ResearchDGX agent

arXiv:2510.08465v2 Announce Type: replace-cross Abstract: Estimating how individual input variables affect the output of a black-box model is a central task in explainable machine learning. However, e

An explicit operator explains end-to-end computation in the modern neural networks used for sequence and language modeling

ApplicationsDGX agent

arXiv:2604.20595v1 Announce Type: cross Abstract: We establish a mathematical correspondence between state space models, a state-of-the-art architecture for capturing long-range dependencies in data,

Applying multimodal biological foundation models across therapeutics and patient care

ApplicationsDGX agent

In this post, we'll explore how multimodal BioFMs work, showcase real-world applications in drug discovery and clinical development, and contextualize how AWS enables organizations to build and deploy

Automated Description Generation of Cytologic Findings for Lung Cytological Images Using a Pretrained Vision Model and Dual Text Decoders: Preliminary Study

ResearchDGX agent

arXiv:2403.18151v2 Announce Type: replace-cross Abstract: Objective: Cytology plays a crucial role in lung cancer diagnosis. Pulmonary cytology involves cell morphological characterization in the spec

← Previous
1…117118119120121…1009
Next →