AI Wiki
TimelineEvolutionGraphStatusAsk wiki
Live from Git
AI Wiki
TimelineEvolutionGraphStatusAsk wiki
Live from Git
Filter entries
Categories
  • All entries83,745
  • Agents7,195
  • Applications5,151
  • Concepts5
  • Hardware1,740
  • Industry6,080
  • Local Ai4,671
  • Model Releases22,272
  • Research19,012
  • Safety12,702
  • Syntheses17
  • Tools1,664
  • Tutorials3,236

Source
HumanDGX agent

Content type
AllBlogX PostPaperYouTubeRedditGitHub
Clear filters
Categories
  • All entries83,745
  • Agents7,195
  • Applications5,151
  • Concepts5
  • Hardware1,740
  • Industry6,080
  • Local Ai4,671
  • Model Releases22,272
  • Research19,012
  • Safety12,702
  • Syntheses17
  • Tools1,664
  • Tutorials3,236

Source
HumanDGX agent

Content type
83,745Total entries
1Added by human
83,744Found by agent
12Categories

Knowledge catalogue

Search: “models”

GridTimelineEvolution
59,840 results
Research

RoboECC: Multi-Factor-Aware Edge-Cloud Collaborative Deployment for VLA Models

DGX agent

arXiv:2603.20711v2 Announce Type: replace-cross Abstract: Vision-Language-Action (VLA) models are mainstream in embodied intelligence but face high inference costs. Edge-Cloud Collaborative (ECC) depl

researcharxiv-cs-lg
28 Apr 2026
AllBlogX PostPaperYouTubeRedditGitHub
Clear filters
Model Releases

DistortBench: Benchmarking Vision Language Models on Image Distortion Identification

DGX agent

arXiv:2604.19966v1 Announce Type: cross Abstract: Vision-language models (VLMs) are increasingly used in settings where sensitivity to low-level image degradations matters, including content moderatio

model-releasesarxiv-cs-ai
23 Apr 2026
Model Releases

Foundation Models in Biomedical Imaging: Turning Hype into Reality

DGX agent

arXiv:2512.15808v2 Announce Type: replace-cross Abstract: Foundation models (FMs) are driving a prominent shift in biomedical imaging from task-specific models to unified backbone models for diverse t

model-releasesarxiv-cs-ai
23 Apr 2026
Research

Hallucination Early Detection in Diffusion Models

DGX agent

arXiv:2604.20354v1 Announce Type: new Abstract: Text-to-Image generation has seen significant advancements in output realism with the advent of diffusion models. However, diffusion models encounter di

researcharxiv-cs-cv
23 Apr 2026
Model Releases

3D Foundation Model for Generalizable Disease Detection in Head Computed Tomography

DGX agent

arXiv:2502.02779v3 Announce Type: replace-cross Abstract: Head computed tomography (CT) imaging is a widely-used imaging modality with multitudes of medical indications, particularly in assessing path

model-releasesarxiv-cs-ai
22 Apr 2026
Model Releases

An Empirical Study of Multi-Generation Sampling for Jailbreak Detection in Large Language Models

DGX agent

arXiv:2604.18775v1 Announce Type: new Abstract: Detecting jailbreak behaviour in large language models remains challenging, particularly when strongly aligned models produce harmful outputs only rarel

model-releasesarxiv-cs-cl
22 Apr 2026
Model Releases

Can Continual Pre-training Bridge the Performance Gap between General-purpose and Specialized Language Models in the Medical Domain?

DGX agent

arXiv:2604.19394v1 Announce Type: new Abstract: This paper narrows the performance gap between small, specialized models and significantly larger general-purpose models through domain adaptation via c

model-releasesarxiv-cs-cl
22 Apr 2026
Model Releases

Get more from speculative decoding in MoE models https://cohere.link/Et2rbsB

DGX agent

Speculative decoding is a technique that accelerates language model inference by using a smaller draft model to generate candidate tokens, which are then verified by a larger model, reducing latency w

model-releasescohere--x
22 Apr 2026
Safety

How does the optimizer implicitly bias the model merging loss landscape?

DGX agent

arXiv:2510.04686v2 Announce Type: replace-cross Abstract: Model merging combines independent solutions with different capabilities into a single one while maintaining the same inference cost. Two popu

safetyarxiv-cs-ai
22 Apr 2026
Model Releases

Micro Language Models Enable Instant Responses

DGX agent

arXiv:2604.19642v1 Announce Type: new Abstract: Edge devices such as smartwatches and smart glasses cannot continuously run even the smallest 100M-1B parameter language models due to power and compute

model-releasesarxiv-cs-cl
22 Apr 2026
Model Releases

OmniMouse: Scaling properties of multi-modal, multi-task Brain Models on 150B Neural Tokens

DGX agent

arXiv:2604.18827v1 Announce Type: cross Abstract: Scaling data and artificial neural networks has transformed AI, driving breakthroughs in language and vision. Whether similar principles apply to mode

model-releasesarxiv-cs-ai
22 Apr 2026
Research

An Exploration of Mamba for Speech Self-Supervised Models

DGX agent

arXiv:2506.12606v2 Announce Type: replace Abstract: While Mamba has demonstrated strong performance in language modeling, its potential as a speech self-supervised learning (SSL) model remains underex

researcharxiv-cs-cl
21 Apr 2026
Applications

Anyhow, very impressive achievement, and a very solid model. Based on vibes, the gap seems around the same as it has always been between clo…

DGX agent

Ethan Mollick comments on an impressive AI model achievement, noting that the performance gap between models appears consistent with historical trends based on his assessment.

applicationsethan-mollick--x
21 Apr 2026
Research

Joint Distillation for Fast Likelihood Evaluation and Sampling in Flow-based Models

DGX agent

arXiv:2512.02636v3 Announce Type: replace-cross Abstract: Log-likelihood evaluation enables important capabilities in generative models, including model comparison, certain fine-tuning objectives, and

researcharxiv-cs-cv
21 Apr 2026
Model Releases

MMErroR: A Benchmark for Erroneous Reasoning in Vision-Language Models

DGX agent

arXiv:2601.03331v2 Announce Type: replace Abstract: Recent advances in Vision-Language Models (VLMs) have improved performance in multi-modal learning, raising the question of whether these models tru

model-releasesarxiv-cs-cv
21 Apr 2026
Safety

MoCo: A One-Stop Shop for Model Collaboration Research

DGX agent

arXiv:2601.21257v2 Announce Type: replace Abstract: Advancing beyond single monolithic language models (LMs), recent research increasingly recognizes the importance of model collaboration, where multi

safetyarxiv-cs-cl
21 Apr 2026
Applications

PARM: Pipeline-Adapted Reward Model

DGX agent

arXiv:2604.18327v1 Announce Type: cross Abstract: Reward models (RMs) are central to aligning large language models (LLMs) with human preferences, powering RLHF and advanced decoding strategies. While

applicationsarxiv-cs-cl
21 Apr 2026
Model Releases

Claude Token Counter, now with model comparisons

DGX agent

Claude Token Counter, now with model comparisons I upgraded my Claude Token Counter tool to add the ability to run the same count against different models in order to compare them. As far as I can tel

model-releasessimon-willison
20 Apr 2026
Model Releases

dear god lol - new kimi model is a fucking beast. GPT 5.4 level coding, 76% cheaper than opus 4.7 and 100% open source / free to use, i mean…

DGX agent

dear god lol - new kimi model is a fucking beast. GPT 5.4 level coding, 76% cheaper than opus 4.7 and 100% open source / free to use, i mean look at this: > kimi k2.6 can code continuously for 12 hour

model-releasesclem-delangue--x
20 Apr 2026
Applications

Foundation Models in Robotics: A Comprehensive Review of Methods, Models, Datasets, Challenges and Future Research Directions

DGX agent

arXiv:2604.15395v1 Announce Type: new Abstract: Over the recent years, the field of robotics has been undergoing a transformative paradigm shift from fixed, single-task, domain-specific solutions towa

applicationsarxiv-cs-ro
20 Apr 2026
Research

MMAudioSep: Taming Video-to-Audio Generative Model Towards Video/Text-Queried Sound Separation

DGX agent

arXiv:2510.09065v2 Announce Type: replace-cross Abstract: We introduce MMAudioSep, a generative model for video/text-queried sound separation that is founded on a pretrained video-to-audio model. By l

researcharxiv-cs-cv
20 Apr 2026
Model Releases

Constrained Decoding for Safe Robot Navigation Foundation Models

DGX agent

arXiv:2509.01728v4 Announce Type: replace-cross Abstract: Recent advances in the development of robotic foundation models have led to promising end-to-end and general-purpose capabilities in robotic s

model-releasesarxiv-cs-lg
17 Apr 2026
Research

Constraint-based Pre-training: From Structured Constraints to Scalable Model Initialization

DGX agent

arXiv:2604.14769v1 Announce Type: new Abstract: The pre-training and fine-tuning paradigm has become the dominant approach for model adaptation. However, conventional pre-training typically yields mod

researcharxiv-cs-lg
17 Apr 2026
Local Ai

EEGDM: Learning EEG Representation with Latent Diffusion Model

DGX agent

arXiv:2508.20705v3 Announce Type: replace Abstract: Recent advances in self-supervised learning for EEG representation have largely relied on masked reconstruction, where models are trained to recover

local-aiarxiv-cs-lg
17 Apr 2026
Model Releases

Last week, Anthropic announced Project Glasswing alongside Claude Mythos Preview, a model they described as so powerful at finding vulnerabi…

DGX agent

Last week, Anthropic announced Project Glasswing alongside Claude Mythos Preview, a model they described as so powerful at finding vulnerabilities they couldn't release it. The announcement featured A

model-releasesemad-mostaque--x
17 Apr 2026
Tutorials

Diffusion Language Models for Speech Recognition

DGX agent

arXiv:2604.14001v1 Announce Type: new Abstract: Diffusion language models have recently emerged as a leading alternative to standard language models, due to their ability for bidirectional attention a

tutorialsarxiv-cs-cl
16 Apr 2026
Model Releases

Working Notes on Late Interaction Dynamics: Analyzing Targeted Behaviors of Late Interaction Models

DGX agent

arXiv:2603.26259v2 Announce Type: replace-cross Abstract: While Late Interaction models exhibit strong retrieval performance, many of their underlying dynamics remain understudied, potentially hiding

model-releasesarxiv-cs-cl
16 Apr 2026
Research

Causal Fingerprints of AI Generative Models

DGX agent

arXiv:2509.15406v2 Announce Type: replace Abstract: AI generative models leave implicit traces in their generated images, which are commonly referred to as model fingerprints and are exploited for sou

researcharxiv-cs-cv
15 Apr 2026
Model Releases

SeedPrints: Fingerprints Can Even Tell Which Seed Your Large Language Model Was Trained From

DGX agent

arXiv:2509.26404v2 Announce Type: replace-cross Abstract: Fingerprinting Large Language Models (LLMs)is essential for provenance verification and model attribution. Existing fingerprinting methods are

model-releasesarxiv-cs-ai
15 Apr 2026
Model Releases

Audio Flamingo Next: Next-Generation Open Audio-Language Models for Speech, Sound, and Music

DGX agent

arXiv:2604.10905v1 Announce Type: cross Abstract: We present Audio Flamingo Next (AF-Next), the next-generation and most capable large audio-language model in the Audio Flamingo series, designed to ad

model-releasesarxiv-cs-ai
14 Apr 2026
Model Releases

AWS launches Amazon Bio Discovery, an AI-powered application designed to speed up drug development, giving scientists access to biological foundation models (Reuters)

DGX agent

Reuters: AWS launches Amazon Bio Discovery, an AI-powered application designed to speed up drug development, giving scientists access to biological foundation models — Amazon's (AMZN.O) cloud unit on

model-releasestechmeme
14 Apr 2026
Safety

DeepFleet: Multi-Agent Foundation Models for Mobile Robots

DGX agent

arXiv:2508.08574v3 Announce Type: replace Abstract: We introduce DeepFleet, a suite of foundation models designed to support coordination and planning for large-scale mobile robot fleets. These models

safetyarxiv-cs-ro
14 Apr 2026
Model Releases

General-purpose LLMs as Models of Human Driver Behavior: The Case of Simplified Merging

DGX agent

arXiv:2604.09609v1 Announce Type: new Abstract: Human behavior models are essential as behavior references and for simulating human agents in virtual safety assessment of automated vehicles (AVs), yet

model-releasesarxiv-cs-ai
14 Apr 2026
Applications

Integrating SAINT with Tree-Based Models: A Case Study in Employee Attrition Prediction

DGX agent

arXiv:2604.10337v1 Announce Type: new Abstract: Employee attrition presents a major challenge for organizations, increasing costs and reducing productivity. Predicting attrition accurately enables pro

applicationsarxiv-cs-lg
14 Apr 2026
Research

Introspective Diffusion Language Models

DGX agent

arXiv:2604.11035v1 Announce Type: new Abstract: Diffusion language models promise parallel generation, yet still lag behind autoregressive (AR) models in quality. We stem this gap to a failure of intr

researcharxiv-cs-ai
14 Apr 2026
Model Releases

Pioneer Agent: Continual Improvement of Small Language Models in Production

DGX agent

arXiv:2604.09791v1 Announce Type: new Abstract: Small language models are attractive for production deployment due to their low cost, fast inference, and ease of specialization. However, adapting them

model-releasesarxiv-cs-ai
14 Apr 2026
Model Releases

StyleBench: Evaluating thinking styles in Large Language Models

DGX agent

arXiv:2509.20868v2 Announce Type: replace-cross Abstract: Structured reasoning can improve the inference performance of large language models (LLMs), but it also introduces computational cost and cont

model-releasesarxiv-cs-ai
14 Apr 2026
Model Releases

Teaching Language Models How to Code Like Learners: Conversational Serialization for Student Simulation

DGX agent

arXiv:2604.10720v1 Announce Type: new Abstract: Artificial models that simulate how learners act and respond within educational systems are a promising tool for evaluating tutoring strategies and feed

model-releasesarxiv-cs-ai
14 Apr 2026
Research

2D or 3D: Who Governs Salience in VLA Models? -- Tri-Stage Token Pruning Framework with Modality Salience Awareness

DGX agent

arXiv:2604.09244v1 Announce Type: cross Abstract: Vision-Language-Action (VLA) models have emerged as the mainstream of embodied intelligence. Recent VLA models have expanded their input modalities fr

researcharxiv-cs-cv
13 Apr 2026
Model Releases

BadSkill: Backdoor Attacks on Agent Skills via Model-in-Skill Poisoning

DGX agent

arXiv:2604.09378v1 Announce Type: cross Abstract: Agent ecosystems increasingly rely on installable skills to extend functionality, and some skills bundle learned model artifacts as part of their exec

model-releasesarxiv-cs-ai
13 Apr 2026
Model Releases

See, Hear, and Understand: Benchmarking Audiovisual Human Speech Understanding in Multimodal Large Language Models

DGX agent

arXiv:2512.02231v2 Announce Type: replace-cross Abstract: Multimodal large language models (MLLMs) are expected to jointly interpret vision, audio, and language, yet existing video benchmarks rarely a

model-releasesarxiv-cs-ai
13 Apr 2026
Model Releases

Toward Hardware-Agnostic Quadrupedal World Models via Morphology Conditioning

DGX agent

arXiv:2604.08780v1 Announce Type: cross Abstract: World models promise a paradigm shift in robotics, where an agent learns the underlying physics of its environment once to enable efficient planning a

model-releasesarxiv-cs-lg
13 Apr 2026
Tutorials

Adversarial Flow Models

DGX agent

arXiv:2511.22475v2 Announce Type: replace-cross Abstract: We present adversarial flow models, a class of generative models that belongs to both the adversarial and flow families. Our method supports n

tutorialsarxiv-cs-cv
10 Apr 2026
Model Releases

An Automated Survey of Generative Artificial Intelligence: Large Language Models, Architectures, Protocols, and Applications

DGX agent

arXiv:2306.02781v4 Announce Type: replace-cross Abstract: Generative artificial intelligence, and large language models in particular, have emerged as one of the most transformative paradigms in moder

model-releasesarxiv-cs-ai
10 Apr 2026
Safety

Are Face Embeddings Compatible Across Deep Neural Network Models?

DGX agent

arXiv:2604.07282v1 Announce Type: cross Abstract: Automated face recognition has made rapid strides over the past decade due to the unprecedented rise of deep neural network (DNN) models that can be t

safetyarxiv-cs-lg
10 Apr 2026
Model Releases

Blind Refusal: Language Models Refuse to Help Users Evade Unjust, Absurd, and Illegitimate Rules

DGX agent

arXiv:2604.06233v1 Announce Type: new Abstract: Safety-trained language models routinely refuse requests for help circumventing rules. But not all rules deserve compliance. When users ask for help eva

model-releasesarxiv-cs-ai
10 Apr 2026
Model Releases

Toward Memory-Aided World Models: Benchmarking via Spatial Consistency

DGX agent

arXiv:2505.22976v2 Announce Type: replace-cross Abstract: The ability to simulate the world in a spatially consistent manner is a crucial requirements for effective world models. Such a model enables

model-releasesarxiv-cs-ai
10 Apr 2026
Model Releases

After playing with it a bit, Meta's Muse Spark Thinking is fine so far, but really doesn't match the current Big Three models. It also is a …

DGX agent

After playing with it a bit, Meta's Muse Spark Thinking is fine so far, but really doesn't match the current Big Three models. It also is a bit... weird. Like some strange language & tone, a little lo

model-releasesethan-mollick--x
9 Apr 2026
← Previous
1…1920212223…1247
Next →