AI Wiki
TimelineEvolutionGraphStatusAsk wiki
Live from Git
AI Wiki
TimelineEvolutionGraphStatusAsk wiki
Live from Git
Filter entries
Categories
  • All entries91,060
  • Agents7,763
  • Applications5,542
  • Concepts5
  • Hardware1,932
  • Industry6,210
  • Local Ai5,103
  • Model Releases24,798
  • Research20,784
  • Safety13,745
  • Syntheses17
  • Tools1,680
  • Tutorials3,481

Source
HumanDGX agent

Content type
AllBlogX PostPaperYouTubeRedditGitHub
Clear filters
Categories
  • All entries91,060
  • Agents7,763
  • Applications5,542
  • Concepts5
  • Hardware1,932
  • Industry6,210
  • Local Ai5,103
  • Model Releases24,798
  • Research20,784
  • Safety13,745
  • Syntheses17
  • Tools1,680
  • Tutorials3,481

Source
HumanDGX agent

Content type
91,060Total entries
1Added by human
91,059Found by agent
12Categories

Knowledge catalogue

Search: “models”

GridTimelineEvolution
65,793 results
Model Releases

Beyond Benchmark Islands: Toward Representative Trustworthiness Evaluation for Agentic AI

DGX agent

arXiv:2603.14987v2 Announce Type: replace Abstract: Agentic AI systems increasingly act through tool-augmented, multi-step workflows whose failures (unsafe tool use, unauthorised actions, social harm)

model-releasesarxiv-cs-cl
22 May 2026
AllBlogX PostPaperYouTubeRedditGitHub
Clear filters
Model Releases

Blind Spots in the Guard: How Domain-Camouflaged Injection Attacks Evade Detection in Multi-Agent LLM Systems

DGX agent

arXiv:2605.22001v1 Announce Type: cross Abstract: Injection detectors deployed to protect LLM agents are calibrated on static, template-based payloads that announce themselves as override directives.

model-releasesarxiv-cs-cl
22 May 2026
Model Releases

Cross-Lingual Consensus: Aligning Multilingual Cultural Knowledge via Multilingual Self-Consistency

DGX agent

arXiv:2605.22137v1 Announce Type: new Abstract: Although Large Language Models (LLMs) demonstrate strong capabilities across various tasks, they exhibit significant performance discrepancies across la

model-releasesarxiv-cs-cl
22 May 2026
Model Releases

DeepSeek’s New AI Is A Game Changer

DGX agent

DeepSeek released two advanced AI models—R1 and V3—with R1 specializing in complex reasoning and V3 designed for large-scale language processing. Notably, DeepSeek developed these high-performing mode

model-releasestwo-minute-papers
22 May 2026
Research

From TF-IDF to Transformers: A Comparative and Ensemble Approach to Sentiment Classification

DGX agent

arXiv:2605.22003v1 Announce Type: new Abstract: Sentiment analysis, also referred to as opinion mining, primarily tries to extract opinion from any text-based data. In the context of movie reviews and

researcharxiv-cs-cl
22 May 2026
Model Releases

GHI: Graphormer over Conditioned Hypergraph Incidence for Aspect-Based Sentiment Analysis

DGX agent

arXiv:2605.22228v1 Announce Type: new Abstract: Aspect-based sentiment analysis (ABSA) requires models to bind sentiment evidence to the correct aspect, making it a natural testbed for fine-grained st

model-releasesarxiv-cs-cl
22 May 2026
Model Releases

H-Flow: Self-supervised Human Scene Flow via Physics-inspired Joint Multi-modal Learning

DGX agent

arXiv:2605.22629v1 Announce Type: new Abstract: Parametric human models capture global pose but cannot represent the non-rigid surface dynamics of clothing and soft tissue. Generic scene flow estimate

model-releasesarxiv-cs-cv
22 May 2026
Model Releases

InteractScience: Programmatic and Visually-Grounded Evaluation of Interactive Scientific Demonstration Code Generation

DGX agent

arXiv:2510.09724v2 Announce Type: replace-cross Abstract: Large Language Models (LLMs) are increasingly capable of generating complete applications from natural language instructions, creating new opp

model-releasesarxiv-cs-ai
22 May 2026
Model Releases

Learning Spatiotemporal Sensitivity in Video LLMs via Counterfactual Reinforcement Learning

DGX agent

arXiv:2605.21988v1 Announce Type: new Abstract: Video large language models (Video LLMs) achieve strong benchmark accuracy, yet often answer video questions through shortcuts such as single-frame cues

model-releasesarxiv-cs-cv
22 May 2026
Model Releases

LongVT: Incentivizing 'Thinking with Long Videos' via Native Tool Calling

DGX agent

arXiv:2511.20785v3 Announce Type: replace Abstract: Large multimodal models (LMMs) have shown great potential for video reasoning with textual Chain-of-Thought. However, they remain vulnerable to hall

model-releasesarxiv-cs-cv
22 May 2026
Model Releases

M3: Conversational LLMs Simplify Secure Clinical Data Access, Understanding, and Analysis

DGX agent

arXiv:2507.01053v4 Announce Type: replace-cross Abstract: Large-scale clinical databases offer opportunities for medical research, but their complexity creates barriers to effective use. The Medical I

model-releasesarxiv-cs-ai
22 May 2026
Model Releases

MotiMotion: Motion-Controlled Video Generation with Visual Reasoning

DGX agent

arXiv:2605.22818v1 Announce Type: new Abstract: Current motion-controlled image-to-video generation models rigidly follow user-provided trajectories that are often sparse, imprecise, and causally inco

model-releasesarxiv-cs-cv
22 May 2026
Model Releases

NuExtract3 released: open-weight 4B VLM for Markdown, OCR and structured extraction (self-hostable) [P]

DGX agent

NuExtract3 is a unified 4B vision-language reasoning model for document understanding that combines structured information extraction with image-to-Markdown conversion, suitable for OCR and RAG prepro

model-releasesr-machinelearning
22 May 2026
Model Releases

oops! wild update, strongly supports @emollick’s overall take:

DGX agent

oops! wild update, strongly supports @emollick’s overall take: I have to eat crow on this, in light of further information. whatever OpenAI spent on Erdos using a new model, apparently you can get GPT

model-releasesgary-marcus--x
22 May 2026
Model Releases

Perception or Prejudice: Can MLLMs Go Beyond First Impressions of Personality?

DGX agent

arXiv:2605.22109v1 Announce Type: cross Abstract: Multimodal Large Language Models (MLLMs) are increasingly deployed in human-facing roles where personality perception is critical, yet existing benchm

model-releasesarxiv-cs-cv
22 May 2026
Model Releases

RAGCap-Bench: Benchmarking Capabilities of LLMs in Agentic Retrieval Augmented Generation Systems

DGX agent

arXiv:2510.13910v2 Announce Type: replace Abstract: Retrieval-Augmented Generation (RAG) mitigates key limitations of Large Language Models (LLMs)-such as factual errors, outdated knowledge, and hallu

model-releasesarxiv-cs-cl
22 May 2026
Model Releases

SpaceDrive: Infusing Spatial Awareness into VLM-based Autonomous Driving

DGX agent

arXiv:2512.10719v2 Announce Type: replace Abstract: End-to-end autonomous driving methods built on vision language models (VLMs) have undergone rapid development driven by their universal visual under

model-releasesarxiv-cs-cv
22 May 2026
Tutorials

TextTeacher: What Can Language Teach About Images?

DGX agent

arXiv:2605.22098v1 Announce Type: new Abstract: The platonic representation hypothesis suggests that sufficiently large models converge to a shared representation geometry, even across modalities. Mot

tutorialsarxiv-cs-cv
22 May 2026
Research

Token-weighted Direct Preference Optimization with Attention

DGX agent

arXiv:2605.21883v1 Announce Type: new Abstract: Direct Preference Optimization (DPO) aligns Large Language Models with human preferences without the need for a separate reward model. However, DPO trea

researcharxiv-cs-cl
22 May 2026
Model Releases

VGenST-Bench: A Benchmark for Spatio-Temporal Reasoning via Active Video Synthesis

DGX agent

arXiv:2605.22570v1 Announce Type: new Abstract: Spatio-temporal reasoning is a core capability for Multimodal Large Language Models (MLLMs) operating in the real world. As such, evaluating it precisel

model-releasesarxiv-cs-cv
22 May 2026
Model Releases

We are making our discount permanent! 🎉 Enjoy building with DeepSeek-V4-Pro and bring your innovative ideas to life! 🚀

DGX agent

DeepSeek has announced a permanent discount for its DeepSeek-V4-Pro model, encouraging developers to build and innovate with the platform. The announcement was made via social media and emphasizes the

model-releasesjeremy-howard--x
22 May 2026
Model Releases

When Cases Get Rare: A Retrieval Benchmark for Off-Guideline Clinical Question Answering

DGX agent

arXiv:2605.21807v1 Announce Type: new Abstract: Across medical specialties, clinical practice is anchored in evidence-based guidelines that codify best studied diagnostic and treatment pathways. These

model-releasesarxiv-cs-cl
22 May 2026
Model Releases

A Free Lunch in LLM Compression: Revisiting Retraining after Pruning

DGX agent

arXiv:2510.14444v3 Announce Type: replace Abstract: Post-training pruning can substantially reduce LLM inference costs, but it often degrades quality unless the remaining weights are adapted. Since gl

model-releasesarxiv-cs-lg
21 May 2026
Model Releases

Anatomy of Agentic Memory: Taxonomy and Empirical Analysis of Evaluation and System Limitations

DGX agent

arXiv:2602.19320v2 Announce Type: replace Abstract: Agentic memory systems enable large language model (LLM) agents to maintain state across long interactions, supporting long-horizon reasoning and pe

model-releasesarxiv-cs-cl
21 May 2026
Model Releases

Break the context window barrier with Amazon Bedrock AgentCore

DGX agent

In this post, you will learn how to implement Recursive Language Models (RLM) using Amazon Bedrock AgentCore Code Interpreter and the Strands Agents SDK. By the end, you will know how to process docum

model-releasesaws-ml-blog
21 May 2026
Research

Can Conversational XAI Improve User Performance? An Experimental Study

DGX agent

arXiv:2605.20439v1 Announce Type: new Abstract: Explainable AI (XAI) techniques aim to provide insights into predictive models and enhance user performance, yet they often fall short of these expectat

researcharxiv-cs-lg
21 May 2026
Model Releases

Causal Path Alignment: Anchoring the Optimization Trajectory for Controllable In-Parameter Knowledge Editing

DGX agent

arXiv:2506.04042v2 Announce Type: replace Abstract: Knowledge editing is pivotal for efficiently updating the parametric memory of Large Language Models (LLMs), enabling them to function as evolving a

model-releasesarxiv-cs-cl
21 May 2026
Model Releases

Cursor's new Composer 2.5 takes third on the Artificial Analysis Coding Agent Index and is ~10-60x lower cost than the higher-effort Opus 4.…

DGX agent

Cursor's new Composer 2.5 takes third on the Artificial Analysis Coding Agent Index and is ~10-60x lower cost than the higher-effort Opus 4.7 and GPT-5.5 variants above it. This release puts Composer

model-releasesfireworks-ai--x
21 May 2026
Model Releases

Dynamic Video Generation: Shaping Video Generation Across Time and Space

DGX agent

arXiv:2605.21042v1 Announce Type: new Abstract: Diffusion models have achieved impressive performance in video generation, but their iterative denoising process remains computationally expensive due t

model-releasesarxiv-cs-cv
21 May 2026
Model Releases

Exploring Deep Learning and Ultra-Widefield Imaging for Diabetic Retinopathy and Macular Edema

DGX agent

arXiv:2603.08235v2 Announce Type: replace Abstract: Diabetic retinopathy (DR) and diabetic macular edema (DME) are leading causes of preventable blindness among working-age adults. Traditional approac

model-releasesarxiv-cs-cv
21 May 2026
Model Releases

Finding the Correct Visual Evidence Without Forgetting: Mitigating Hallucination in LVLMs via Inter-Layer Visual Attention Discrepancy

DGX agent

arXiv:2605.20965v1 Announce Type: new Abstract: Large Vision-Language Models (LVLMs) have shown remarkable performance on a wide range of vision-language tasks. Despite this progress, they are still p

model-releasesarxiv-cs-cv
21 May 2026
Model Releases

FT-Dojo: Towards Autonomous LLM Fine-Tuning with Language Agents

DGX agent

arXiv:2603.01712v2 Announce Type: replace-cross Abstract: Fine-tuning large language models for vertical domains remains labor-intensive, requiring practitioners to curate data, configure training, an

model-releasesarxiv-cs-lg
21 May 2026
Model Releases

How Much Online RL is Enough? Informative Rollouts for Offline Preference Optimization in RLVR

DGX agent

arXiv:2605.21266v1 Announce Type: new Abstract: Reinforcement Learning from Verifiable Rewards (RLVR) has emerged as a powerful paradigm for reasoning in language models, with GRPO as its primary exam

model-releasesarxiv-cs-lg
21 May 2026
Model Releases

Learning Query-Aware Budget-Tier Routing for Runtime Agent Memory

DGX agent

arXiv:2602.06025v2 Announce Type: replace Abstract: Memory is increasingly central to Large Language Model (LLM) agents operating beyond a single context window, yet most existing systems rely on offl

model-releasesarxiv-cs-cl
21 May 2026
Local Ai

LoCar: Localization-Aware Evaluation of In-Vehicle Assistants through Fine-Grained Sociolinguistic Control

DGX agent

arXiv:2605.21086v1 Announce Type: new Abstract: While Large Language Models (LLMs) are increasingly integrated into in-vehicle conversational systems, identifying the optimal model remains challenging

local-aiarxiv-cs-cl
21 May 2026
Model Releases

Mechanics of Bias and Reasoning: Interpreting the Impact of Chain-of-Thought Prompting on Gender Bias in LLMs

DGX agent

arXiv:2605.20410v1 Announce Type: new Abstract: Large language models (LLMs) are increasingly deployed in socially sensitive settings despite substantial documentation that they encode gender biases.

model-releasesarxiv-cs-cl
21 May 2026
Safety

Mechanistic Interpretability for Learning Assurance of a Vision-Based Landing System

DGX agent

arXiv:2605.20607v1 Announce Type: cross Abstract: EASA's learning-assurance guidance requires data-driven aviation systems to build and monitor their own situation representation, yet for neural netwo

safetyarxiv-cs-cv
21 May 2026
Model Releases

PGC: Peak-Guided Calibration for Generalizable AI-Generated Image Detection

DGX agent

arXiv:2605.21207v1 Announce Type: new Abstract: The rapid evolution of generative AI, from GANs to modern diffusion models, has resulted in increasingly subtle discriminative clues. These fine-grained

model-releasesarxiv-cs-cv
21 May 2026
Model Releases

QwenSafe: Multimodal Content Rating Description Identification via Preference-Aligned VLMs

DGX agent

arXiv:2605.20584v1 Announce Type: new Abstract: Mobile app marketplaces require developers to disclose standardized content rating descriptors (CRDs) to inform users about potentially sensitive or res

model-releasesarxiv-cs-cv
21 May 2026
Model Releases

RankE: End-to-End Post-Training for Discrete Text-to-Image Generation with Decoder Co-Evolution

DGX agent

arXiv:2605.21195v1 Announce Type: new Abstract: Discrete autoregressive (AR) text-to-image (T2I) models pair a VQ tokenizer with an AR policy, and current post-training pipelines optimize only the pol

model-releasesarxiv-cs-cv
21 May 2026
Model Releases

Refining and Reusing Annotation Guidelines for LLM Annotation

DGX agent

arXiv:2605.20809v1 Announce Type: new Abstract: While Large Language Models (LLMs) demonstrate remarkable performance on zero-shot annotation tasks, they often struggle with the specialized convention

model-releasesarxiv-cs-cl
21 May 2026
Model Releases

SOLAR: A Self-Optimizing Open-Ended Autonomous Agent for Lifelong Learning and Continual Adaptation

DGX agent

arXiv:2605.20189v1 Announce Type: cross Abstract: Despite the remarkable success of large language models (LLMs), they still face bottlenecks while deploying in dynamic, real-world settings with prima

model-releasesarxiv-cs-lg
21 May 2026
Local Ai

Sustainability Is Not Linear: Quantifying Performance, Energy, and Privacy Trade-offs in On-Device Intelligence

DGX agent

arXiv:2603.26603v2 Announce Type: replace-cross Abstract: The migration of Large Language Models (LLMs) from cloud clusters to edge devices promises enhanced privacy and offline accessibility, but thi

local-aiarxiv-cs-lg
21 May 2026
Model Releases

TASTE: A Designer-Annotated Multi-Dimensional Preference Dataset for AI-Generated Graphic Design

DGX agent

arXiv:2605.20731v1 Announce Type: new Abstract: Text-to-image models produce graphic design at production scale, but their supervision comes from photo-style preference data with a single overall verd

model-releasesarxiv-cs-cv
21 May 2026
Model Releases

Terminal-World: Scaling Terminal-Agent Environments via Agent Skills

DGX agent

arXiv:2605.20876v1 Announce Type: new Abstract: Terminal agents extend Large Language Models with the ability to execute tasks directly in command-line environments, but their progress is bottlenecked

model-releasesarxiv-cs-cl
21 May 2026
Research

TreeText-CTS: Compact, Source-Traceable Tree-Path Evidence for Irregular Clinical Time-Series Prediction

DGX agent

arXiv:2605.20292v1 Announce Type: new Abstract: Numerical time-series models can effectively process irregular electronic health record (EHR) trajectories, but they do not naturally expose the measure

researcharxiv-cs-lg
21 May 2026
Safety

What Semantics Survive the Connector? Diagnosing VLM-to-DiT Alignment in Video Editing

DGX agent

arXiv:2605.20795v1 Announce Type: new Abstract: Flow matching based video generative models have been increasingly relying on prepended Vision-Language Models (VLMs) to handle complex, instruction-bas

safetyarxiv-cs-cv
21 May 2026
Model Releases

A Bitter Lesson for Data Filtering

DGX agent

arXiv:2605.19407v1 Announce Type: cross Abstract: We investigate data filtering for large model pretraining via new scaling studies that target the high compute, data-scarce regime. In spite of an app

model-releasesarxiv-cs-ai
20 May 2026
← Previous
1…463464465466467…1371
Next →