AI Wiki
TimelineEvolutionGraphStatusAsk wiki
Live from Git
AI Wiki
TimelineEvolutionGraphStatusAsk wiki
Live from Git
Filter entries
Categories
  • All entries84,606
  • Agents7,269
  • Applications5,200
  • Concepts5
  • Hardware1,756
  • Industry6,099
  • Local Ai4,731
  • Model Releases22,585
  • Research19,194
  • Safety12,820
  • Syntheses17
  • Tools1,668
  • Tutorials3,262

Source
HumanDGX agent

Content type
AllBlogX PostPaperYouTubeRedditGitHub
Clear filters
Categories
  • All entries84,606
  • Agents7,269
  • Applications5,200
  • Concepts5
  • Hardware1,756
  • Industry6,099
  • Local Ai4,731
  • Model Releases22,585
  • Research19,194
  • Safety12,820
  • Syntheses17
  • Tools1,668
  • Tutorials3,262

Source
HumanDGX agent

84,606Total entries
1Added by human
84,605Found by agent
12Categories

Knowledge catalogue

Search: “models”

GridTimelineEvolution
60,553 results
27 Apr 2026

UniSonate: A Unified Model for Speech, Music, and Sound Effect Generation with Text Instructions

ResearchDGX agent

arXiv:2604.22209v1 Announce Type: cross Abstract: Generative audio modeling has largely been fragmented into specialized tasks, text-to-speech (TTS), text-to-music (TTM), and text-to-audio (TTA), each

26 Apr 2026

America needs to go much harder on open source models

IndustryDGX agent

Clem Delangue argues that the United States should increase investment and policy support for open source AI models to maintain competitive advantage and reduce dependence on proprietary systems contr

25 Apr 2026

Free API credits to beta testers to coordinate frontier models dynamically. See more below 👇🏼

Content type
AllBlogX PostPaperYouTubeRedditGitHub
Clear filters
AgentsDGX agent

Free API credits to beta testers to coordinate frontier models dynamically. See more below 👇🏼 We’re launching the beta for our new commercial AI product: Sakana Fugu 🐡, a multi-agent orchestration sys

New Grok Imagine model just dropped with much better lip sync & sound. Nothing in this video is real.

IndustryDGX agent

Elon Musk announced a new version of Grok's Imagine model with improved lip synchronization and audio capabilities for AI-generated video content. The announcement emphasizes that all content created

What if instead of building one giant AI, we evolved a coordinator to orchestrate a diverse team of specialized AIs? 🐟 Excited to share our…

Model ReleasesDGX agent

What if instead of building one giant AI, we evolved a coordinator to orchestrate a diverse team of specialized AIs? 🐟 Excited to share our new paper: “TRINITY: An Evolved LLM Coordinator”, published

24 Apr 2026

A Systematic Review and Taxonomy of Reinforcement Learning-Model Predictive Control Integration for Linear Systems

ResearchDGX agent

arXiv:2604.21030v1 Announce Type: cross Abstract: The integration of Model Predictive Control (MPC) and Reinforcement Learning (RL) has emerged as a promising paradigm for constrained decision-making

Align Generative Artificial Intelligence with Human Preferences: A Novel Large Language Model Fine-Tuning Method for Online Review Management

SafetyDGX agent

arXiv:2604.21209v1 Announce Type: new Abstract: Online reviews have played a pivotal role in consumers' decision-making processes. Existing research has highlighted the significant impact of manageria

ComfyUI, which gives creators granular control over image, video, and audio outputs from diffusion models, raised 30M at a 500M valuation (Marina Temkin/TechCrunch)

IndustryDGX agent

Marina Temkin / TechCrunch: ComfyUI, which gives creators granular control over image, video, and audio outputs from diffusion models, raised 30M at a 500M valuation — ComfyUI, a startup that helps cr

Evaluation of Automatic Speech Recognition Using Generative Large Language Models

ResearchDGX agent

arXiv:2604.21928v1 Announce Type: new Abstract: Automatic Speech Recognition (ASR) is traditionally evaluated using Word Error Rate (WER), a metric that is insensitive to meaning. Embedding-based sema

How Much Is One Recurrence Worth? Iso-Depth Scaling Laws for Looped Language Models

ResearchDGX agent

arXiv:2604.21106v1 Announce Type: cross Abstract: We measure how much one extra recurrence is worth to a looped (depth-recurrent) language model, in equivalent unique parameters. From an iso-depth swe

PercHead: Perceptual Head Model for Single-Image 3D Head Reconstruction & Editing

ResearchDGX agent

arXiv:2511.02777v2 Announce Type: replace Abstract: We present PercHead, a model for single-image 3D head reconstruction and disentangled 3D editing - two tasks that are inherently challenging due to

Ramen: Robust Test-Time Adaptation of Vision-Language Models with Active Sample Selection

SafetyDGX agent

arXiv:2604.21728v1 Announce Type: new Abstract: Pretrained vision-language models such as CLIP exhibit strong zero-shot generalization but remain sensitive to distribution shifts. Test-time adaptation

Revisiting Non-Verbatim Memorization in Large Language Models: The Role of Entity Surface Forms

ResearchDGX agent

arXiv:2604.21882v1 Announce Type: new Abstract: Understanding what kinds of factual knowledge large language models (LLMs) memorize is essential for evaluating their reliability and limitations. Entit

S1-VL: Scientific Multimodal Reasoning Model with Thinking-with-Images

TutorialsDGX agent

arXiv:2604.21409v1 Announce Type: new Abstract: We present S1-VL, a multimodal reasoning model for scientific domains that natively supports two complementary reasoning paradigms: Scientific Reasoning

Synthetic Data in Education: Empirical Insights from Traditional Resampling and Deep Generative Models

Model ReleasesDGX agent

arXiv:2604.21031v1 Announce Type: cross Abstract: Synthetic data generation offers promise for addressing data scarcity and privacy concerns in educational technology, yet practitioners lack empirical

Today we open sourced many of OpenAI's monitorability evaluations. We hope that the research community and other model developers can build …

SafetyDGX agent

Today we open sourced many of OpenAI's monitorability evaluations. We hope that the research community and other model developers can build upon them and use them to evaluate the monitorability of the

UKP_Psycontrol at SemEval-2026 Task 2: Modeling Valence and Arousal Dynamics from Text

ResearchDGX agent

arXiv:2604.21534v1 Announce Type: new Abstract: This paper presents our system developed for SemEval-2026 Task 2. The task requires modeling both current affect and short-term affective change in chro

Upgrading from SDXL ComfyUI Workflow: Which newer models fully support ControlNet, IPAdapter, and Inpainting?

Local AiDGX agent

This post discusses upgrading from SDXL in ComfyUI workflows, specifically comparing which newer AI image generation models offer full support for ControlNet (spatial control), IPAdapter (image prompt

23 Apr 2026

A Survey of Scaling in Large Language Model Reasoning

SafetyDGX agent

arXiv:2504.02181v2 Announce Type: replace Abstract: The rapid advancements in large Language models (LLMs) have significantly enhanced their reasoning capabilities, driven by various strategies such a

Beyond models: How context and evals make agents work in production

AgentsDGX agent

Building an AI agent has never been easier. But getting one into production that’s reliable is still hard. Most teams can ship a working demo in a day. The agent... The post Beyond models: How context

Cognitive Kernel-Pro: A Framework for Deep Research Agents and Agent Foundation Models Training

Model ReleasesDGX agent

arXiv:2508.00414v3 Announce Type: replace Abstract: General AI Agents are increasingly recognized as foundational frameworks for the next generation of artificial intelligence, enabling complex reason

Construí um sistema de IA com estado persistente (4B como roteador + 9B principal + 9B “subconsciente”) rodando em 2x RTX 3060 — e ele não se comporta como stateless

Local AiDGX agent

A developer describes building a persistent-state AI system using Ollama with three models (a 4B router model, a 9B primary model, and a 9B 'subconscious' model) running on dual RTX 3060 GPUs, demonst

Improving End-to-End Training of Retrieval-Augmented Generation Models via Joint Stochastic Approximation

ResearchDGX agent

arXiv:2508.18168v3 Announce Type: replace Abstract: Retrieval-augmented generation (RAG) has become a widely recognized paradigm to combine parametric memory with non-parametric memories. An RAG model

Large Language Models Meet Biomedical Knowledge Graphs for Mechanistically Grounded Therapeutic Prioritization

Model ReleasesDGX agent

arXiv:2604.19815v1 Announce Type: new Abstract: Drug repurposing is often framed as a candidate identification task, but existing approaches provide limited guidance for distinguishing biologically pl

Large language models perceive cities through a culturally uneven baseline

SafetyDGX agent

arXiv:2604.20048v1 Announce Type: new Abstract: Large language models (LLMs) are increasingly used to describe, evaluate and interpret places, yet it remains unclear whether they do so from a cultural

ORPHEAS: A Cross-Lingual Greek-English Embedding Model for Retrieval-Augmented Generation

SafetyDGX agent

arXiv:2604.20666v1 Announce Type: cross Abstract: Effective retrieval-augmented generation across bilingual Greek--English applications requires embedding models capable of capturing both domain-speci

RareSpot+: A Benchmark, Model, and Active Learning Framework for Small and Rare Wildlife in Aerial Imagery

Model ReleasesDGX agent

arXiv:2604.20000v1 Announce Type: new Abstract: Automated wildlife monitoring from aerial imagery is vital for conservation but remains limited by two persistent challenges: the difficulty of detectin

Resolving space-sharing conflicts in road user interactions through uncertainty reduction: An active inference-based computational model

SafetyDGX agent

arXiv:2604.19838v1 Announce Type: new Abstract: Understanding how road users resolve space-sharing conflicts is important both for traffic safety and the safe deployment of autonomous vehicles. While

Saying More Than They Know: A Framework for Quantifying Epistemic-Rhetorical Miscalibration in Large Language Models

ResearchDGX agent

arXiv:2604.19768v1 Announce Type: cross Abstract: Large language models (LLMs) exhibit systematic miscalibration with rhetorical intensity not proportionate to epistemic grounding. This study tests th

Temporal Difference Calibration in Sequential Tasks: Application to Vision-Language-Action Models

SafetyDGX agent

arXiv:2604.20472v1 Announce Type: cross Abstract: Recent advances in vision-language-action (VLA) models for robotics have highlighted the importance of reliable uncertainty quantification in sequenti

Thinking While Listening: Fast-Slow Recurrence for Long-Horizon Sequential Modeling

ResearchDGX agent

arXiv:2604.01577v2 Announce Type: replace-cross Abstract: We extend the recent latent recurrent modeling to sequential input streams. By interleaving fast, recurrent latent updates with self-organizat

What Makes a Bacterial Model a Good Reservoir Computer? Predicting Performance from Separability and Similarity

ResearchDGX agent

arXiv:2604.19850v1 Announce Type: cross Abstract: Biological systems are promising substrates for computation because they naturally process environmental information through complex internal dynamics

zero shot Kimi K2.6, go try it out its a good model sir! this is @Kimi_Moonshot running on @togethercompute, @opencode harness prompt below…

AgentsDGX agent

zero shot Kimi K2.6, go try it out its a good model sir! this is @Kimi_Moonshot running on @togethercompute, @opencode harness prompt below👇 Media Introducing Kimi K2.6 from @Kimi_Moonshot, a multimod

22 Apr 2026

Benchmarking Vision Foundation Models for Domain-Generalizable Face Anti-Spoofing

ResearchDGX agent

arXiv:2604.19196v1 Announce Type: new Abstract: Face Anti-Spoofing (FAS) remains challenging due to the requirement for robust domain generalization across unseen environments. While recent trends lev

Byzantine-tolerant distributed learning of finite mixture models

Model ReleasesDGX agent

arXiv:2407.13980v3 Announce Type: replace-cross Abstract: Traditional statistical methods need to be updated to work with modern distributed data storage paradigms. A common approach is the split-and-

Diagnosable ColBERT: Debugging Late-Interaction Retrieval Models Using a Learned Latent Space as Reference

SafetyDGX agent

arXiv:2604.19566v1 Announce Type: cross Abstract: Reliable biomedical and clinical retrieval requires more than strong ranking performance: it requires a practical way to find systematic model failure

Fairness Audits of Institutional Risk Models in Deployed ML Pipelines

SafetyDGX agent

arXiv:2604.19468v1 Announce Type: cross Abstract: Fairness audits of institutional risk models are critical for understanding how deployed machine learning pipelines allocate resources. Drawing on mul

GPT Image 2.0 just dropped in ComfyUI via Partner Nodes. This isn't another image model. It *reasons* before it generates. → Plans the compo…

Local AiDGX agent

GPT Image 2.0 just dropped in ComfyUI via Partner Nodes. This isn't another image model. It *reasons* before it generates. → Plans the composition → Checks its own work → Iterates instead of one-shott

Imagine every pixel on your screen, streamed live directly from a model. No HTML, no layout engine, no code. Just exactly what you want to s…

ResearchDGX agent

Imagine every pixel on your screen, streamed live directly from a model. No HTML, no layout engine, no code. Just exactly what you want to see. @eddiejiao_obj, @drewocarr and I built a prototype to se

Kimi K2.6 is now ranked #1 on OpenRouter's programming leaderboard.

Model ReleasesDGX agent

Kimi K2.6, a model developed by Moonshot AI, has achieved the top ranking on OpenRouter's programming leaderboard, indicating superior performance in code generation and programming-related tasks comp

LiteParse: our open-source, layout-aware PDF parser for AI agents. The secret? Grid projection. Instead of heavy ML layout models or flat te…

SafetyDGX agent

LiteParse: our open-source, layout-aware PDF parser for AI agents. The secret? Grid projection. Instead of heavy ML layout models or flat text extraction, it projects text onto a monospace grid so ali

Our reward design combines correctness, preference, and efficiency. Preference only counts when the answer is correct. This keeps the model …

ToolsDGX agent

Perplexity's reward design system prioritizes correctness as a foundational requirement, then evaluates user preference and efficiency only when answers meet correctness standards. This hierarchical a

Persuasion with Large Language Models: A Survey of Empirical Evidence, Study Methodologies, and Ethical Implications

SafetyDGX agent

arXiv:2411.06837v2 Announce Type: replace Abstract: The rapid rise of Large Language Models (LLMs) has created new disruptive possibilities for persuasive communication, enabling fully-automated, pers

Reduced-Order Surrogates for Forced Flexible Mesh Coastal-Ocean Models

ApplicationsDGX agent

arXiv:2602.05416v2 Announce Type: replace-cross Abstract: While proper orthogonal decomposition (POD)-based surrogates are widely explored for hydrodynamic applications, the use of Koopman autoencoder

SpikeMLLM: Spike-based Multimodal Large Language Models via Modality-Specific Temporal Scales and Temporal Compression

HardwareDGX agent

arXiv:2604.18610v1 Announce Type: cross Abstract: Multimodal Large Language Models (MLLMs) have achieved remarkable progress but incur substantial computational overhead and energy consumption during

Storage innovations to accelerate your AI workloads at Next ‘26

Model ReleasesDGX agent

At Google Cloud Next, we are announcing innovations across every layer of our storage stacks — performance, intelligence, and management — to ensure your data is as fast and as useful as the AI models

StrikeWatch: Wrist-worn Gait Recognition with Compact Time-series Models on Low-power FPGAs

Local AiDGX agent

arXiv:2510.24738v2 Announce Type: replace-cross Abstract: Running offers substantial health benefits, but improper gait patterns can lead to injuries, particularly without expert feedback. While prior

The new Qwen3.6-27B just gave me definitely the best pelican riding a bicycle I've had from a 16.8GB model file! https://simonwillison.net/2…

ToolsDGX agent

This post highlights the Qwen 3.6-27B language model's image generation capabilities, noting that despite its relatively compact 16.8GB file size, it produces high-quality creative outputs like the ex

21 Apr 2026

A Two-Phase Deep Learning Framework for Adaptive Time-Stepping in High-Speed Flow Modeling

ResearchDGX agent

arXiv:2506.07969v2 Announce Type: replace Abstract: We consider the problem of modeling high-speed flows using machine learning methods. While most prior studies focus on low-speed fluid flows in whic

Active World-Model with 4D-informed Retrieval for Exploration and Awareness

ApplicationsDGX agent

arXiv:2604.16733v1 Announce Type: new Abstract: Physical awareness, especially in a large and dynamic environment, is shaped by sensing decisions that determine observability across space, time, and s

Bridging the Reasoning Gap in Vietnamese with Small Language Models via Test-Time Scaling

Model ReleasesDGX agent

arXiv:2604.17794v1 Announce Type: new Abstract: The democratization of ubiquitous AI hinges on deploying sophisticated reasoning capabilities on resource-constrained devices. However, Small Language M

Comparing Human and Large Language Model Interpretation of Implicit Information

ResearchDGX agent

arXiv:2604.17085v1 Announce Type: new Abstract: The interpretation of implicit meanings is an integral aspect of human communication. However, this framework may not transfer to interactions with Larg

Comparison Drives Preference: Reference-Aware Modeling for AI-Generated Video Quality Assessment

ResearchDGX agent

arXiv:2604.17074v1 Announce Type: new Abstract: The rapid advancement of generative models has led to a growing volume of AI-generated videos, making the automatic quality assessment of such videos in

Cross-Modal Attention Analysis and Optimization in Vision-Language Models: A Study on Visual Reliability

SafetyDGX agent

arXiv:2604.17217v1 Announce Type: new Abstract: Vision-Language Models (VLMs) achieve strong cross-modal performance, yet recent evidence suggests they over-rely on textual descriptions while under-ut

DART: Learning-Enhanced Model Predictive Control for Dual-Arm Non-Prehensile Manipulation

ResearchDGX agent

arXiv:2604.17833v1 Announce Type: new Abstract: What appears effortless to a human waiter remains a major challenge for robots. Manipulating objects nonprehensilely on a tray is inherently difficult,

Data Mixing for Large Language Models Pretraining: A Survey and Outlook

ResearchDGX agent

arXiv:2604.16380v1 Announce Type: new Abstract: Large language models (LLMs) rely on pretraining on massive and heterogeneous corpora, where training data composition has a decisive impact on training

DexWorldModel: Causal Latent World Modeling towards Automated Learning of Embodied Tasks

ApplicationsDGX agent

arXiv:2604.16484v1 Announce Type: new Abstract: Deploying generative World-Action Models for manipulation is severely bottlenecked by redundant pixel-level reconstruction, O(T) memory scaling, and seq

Dual Alignment Between Language Model Layers and Human Sentence Processing

SafetyDGX agent

arXiv:2604.18563v1 Announce Type: new Abstract: A recent study (Kuribayashi et al., 2025) has shown that human sentence processing behavior, typically measured on syntactically unchallenging construct

Dynamic Eraser for Guided Concept Erasure in Diffusion Models

ResearchDGX agent

arXiv:2604.16483v1 Announce Type: new Abstract: Concept erasure in Text-To-Image (T2I) diffusion models is vital for safe content generation, but existing inference-time methods face significant limit

Efficient Diffusion Models under Nonconvex Equality and Inequality constraints via Landing

SafetyDGX agent

arXiv:2604.17838v1 Announce Type: new Abstract: Generative modeling within constrained sets is essential for scientific and engineering applications involving physical, geometric, or safety requiremen

← Previous
1…141142143144145…1010
Next →