AI Wiki
TimelineEvolutionGraphStatusAsk wiki
Live from Git
AI Wiki
TimelineEvolutionGraphStatusAsk wiki
Live from Git
Filter entries
Categories
  • All entries84,648
  • Agents7,273
  • Applications5,201
  • Concepts5
  • Hardware1,758
  • Industry6,104
  • Local Ai4,732
  • Model Releases22,612
  • Research19,194
  • Safety12,821
  • Syntheses17
  • Tools1,669
  • Tutorials3,262

Source
HumanDGX agent

Content type
AllBlogX PostPaperYouTubeRedditGitHub
Clear filters
Categories
  • All entries84,648
  • Agents7,273
  • Applications5,201
  • Concepts5
  • Hardware1,758
  • Industry6,104
  • Local Ai4,732
  • Model Releases22,612
  • Research19,194
  • Safety12,821
  • Syntheses17
  • Tools1,669
  • Tutorials3,262

Source
HumanDGX agent

Content type
84,648Total entries
1Added by human
84,647Found by agent
12Categories

Knowledge catalogue

Search: “models”

GridTimelineEvolution
60,587 results
Model Releases

MetaGAI: A Large-Scale and High-Quality Benchmark for Generative AI Model and Data Card Generation

DGX agent

arXiv:2604.23539v1 Announce Type: new Abstract: The rapid proliferation of Generative AI necessitates rigorous documentation standards for transparency and governance. However, manual creation of Mode

model-releasesarxiv-cs-ai
28 Apr 2026
AllBlogX PostPaperYouTubeRedditGitHub
Clear filters
Research

Modeling Induced Pleasure through Cognitive Appraisal Prediction via Multimodal Fusion

DGX agent

arXiv:2604.23753v1 Announce Type: new Abstract: Multimodal affective computing analyzes user-generated social media content to predict emotional states. However, a critical gap remains in understandin

researcharxiv-cs-ai
28 Apr 2026
Applications

Modular Sensory Stream for Integrating Physical Feedback in Vision-Language-Action Models

DGX agent

arXiv:2604.23272v1 Announce Type: new Abstract: Humans understand and interact with the real world by relying on diverse physical feedback beyond visual perception. Motivated by this, recent approache

applicationsarxiv-cs-ro
28 Apr 2026
Model Releases

OntoLogX: Ontology-Guided Knowledge Graph Extraction from Cybersecurity Logs with Large Language Models

DGX agent

arXiv:2510.01409v2 Announce Type: replace Abstract: System logs represent a valuable source of Cyber Threat Intelligence (CTI), capturing attacker behaviors, exploited vulnerabilities, and traces of m

model-releasesarxiv-cs-ai
28 Apr 2026
Safety

Reflective Flow Sampling Enhancement

DGX agent

arXiv:2603.06165v2 Announce Type: replace-cross Abstract: The growing demand for text-to-image generation has led to rapid advances in generative modeling. Recently, text-to-image diffusion models tra

safetyarxiv-cs-ai
28 Apr 2026
Safety

RL Token: Bootstrapping Online RL with Vision-Language-Action Models

DGX agent

arXiv:2604.23073v1 Announce Type: new Abstract: Vision-language-action (VLA) models can learn to perform diverse manipulation skills 'out of the box,' but achieving the precision and speed that real-w

safetyarxiv-cs-lg
28 Apr 2026
Agents

SCRIBE: Structured Mid-Level Supervision for Tool-Using Language Models

DGX agent

arXiv:2601.03555v2 Announce Type: replace Abstract: Training reliable tool-augmented agents remains a significant challenge, largely due to the difficulty of credit assignment in multi-step reasoning.

agentsarxiv-cs-ai
28 Apr 2026
Model Releases

Secure On-Premise Deployment of Open-Weights Large Language Models in Radiology: An Isolation-First Architecture with Prospective Pilot Evaluation

DGX agent

arXiv:2604.22768v1 Announce Type: cross Abstract: Purpose: To design, implement, evaluate, and report on the regulatory requirements of a self-hosted LLM infrastructure for radiology adhering to the p

model-releasesarxiv-cs-cl
28 Apr 2026
Safety

Seer: Language Instructed Video Prediction with Latent Diffusion Models

DGX agent

arXiv:2303.14897v4 Announce Type: replace Abstract: Imagining the future trajectory is the key for robots to make sound planning and successfully reach their goals. Therefore, text-conditioned video p

safetyarxiv-cs-cv
28 Apr 2026
Model Releases

Stress-Testing Emotional Support Models: Moving from Homogeneous to Diverse Help Seekers

DGX agent

arXiv:2601.07698v2 Announce Type: replace Abstract: As emotional support chatbots have recently gained significant traction across both research and industry, a common evaluation strategy has emerged:

model-releasesarxiv-cs-cl
28 Apr 2026
Model Releases

TexOCR: Advancing Document OCR Models for Compilable Page-to-LaTeX Reconstruction

DGX agent

arXiv:2604.22880v1 Announce Type: new Abstract: Existing document OCR largely targets plain text or Markdown, discarding the structural and executable properties that make LaTeX essential for scientif

model-releasesarxiv-cs-cl
28 Apr 2026
Agents

The Chameleon's Limit: Investigating Persona Collapse and Homogenization in Large Language Models

DGX agent

arXiv:2604.24698v1 Announce Type: new Abstract: Applications based on large language models (LLMs), such as multi-agent simulations, require population diversity among agents. We identify a pervasive

agentsarxiv-cs-cl
28 Apr 2026
Model Releases

The Surprising Effectiveness of Membership Inference with Simple N-Gram Coverage

DGX agent

arXiv:2508.09603v2 Announce Type: replace Abstract: Membership inference attacks serves as useful tool for fair use of language models, such as detecting potential copyright infringement and auditing

model-releasesarxiv-cs-cl
28 Apr 2026
Applications

Training Machine Learning Models on Encrypted Data: A Privacy-Preserving Framework using Homomorphic Encryption

DGX agent

arXiv:2604.23245v1 Announce Type: cross Abstract: The use of Machine Learning (ML) for data-driven decision-making often relies on access to sensitive datasets, which introduces privacy challenges. Tr

applicationsarxiv-cs-ai
28 Apr 2026
Applications

Variational Grey-Box Dynamics Matching

DGX agent

arXiv:2602.17477v3 Announce Type: replace Abstract: Deep generative models such as flow matching and diffusion models have shown great potential in learning complex distributions and dynamical systems

applicationsarxiv-cs-lg
28 Apr 2026
Research

VS-DDPM: Efficient Low-Cost Diffusion Model for Medical Modality Translation

DGX agent

arXiv:2604.22942v1 Announce Type: cross Abstract: Diffusion models produce high-quality synthetic data but suffer from slow inference. We propose 3D Variable-Step Denoising Diffusion Probabilistic Mod

researcharxiv-cs-ai
28 Apr 2026
Research

Weakly Supervised Multicenter Nancy Index Scoring in Ulcerative Colitis Using Foundation Models

DGX agent

arXiv:2604.23706v1 Announce Type: new Abstract: Histologic assessment of ulcerative colitis (UC) activity is an important endpoint in clinical trials and routine care, but manual grading with indices

researcharxiv-cs-cv
28 Apr 2026
Applications

When Silence Matters: The Impact of Irrelevant Audio on Text Reasoning in Large Audio-Language Models

DGX agent

arXiv:2510.00626v3 Announce Type: replace-cross Abstract: Large audio-language models (LALMs) unify speech and text processing, but their robustness in noisy real-world settings remains underexplored.

applicationsarxiv-cs-cl
28 Apr 2026
Safety

Clutter-Robust Vision-Language-Action Models through Object-Centric and Geometry Grounding

DGX agent

arXiv:2512.22519v2 Announce Type: replace Abstract: Recent Vision-Language-Action (VLA) models have made impressive progress toward general-purpose robotic manipulation by post-training large Vision-L

safetyarxiv-cs-ro
27 Apr 2026
Agents

DVGT-2: Vision-Geometry-Action Model for Autonomous Driving at Scale

DGX agent

arXiv:2604.00813v3 Announce Type: replace-cross Abstract: End-to-end autonomous driving has evolved from the conventional paradigm based on sparse perception into vision-language-action (VLA) models,

agentsarxiv-cs-ai
27 Apr 2026
Research

Initial results of the Digital Consciousness Model

DGX agent

arXiv:2601.17060v2 Announce Type: replace-cross Abstract: Artificially intelligent systems have become remarkably sophisticated. They hold conversations, write essays, and seem to understand context i

researcharxiv-cs-ai
27 Apr 2026
Model Releases

LLMPhy: Parameter-Identifiable Physical Reasoning Combining Large Language Models and Physics Engines

DGX agent

arXiv:2411.08027v3 Announce Type: replace-cross Abstract: Most learning-based approaches to complex physical reasoning sidestep the crucial problem of parameter identification (e.g., mass, friction) t

model-releasesarxiv-cs-ai
27 Apr 2026
Research

Preference Heads in Large Language Models: A Mechanistic Framework for Interpretable Personalization

DGX agent

arXiv:2604.22345v1 Announce Type: new Abstract: Large Language Models (LLMs) exhibit strong implicit personalization ability, yet most existing approaches treat this behavior as a black box, relying o

researcharxiv-cs-cl
27 Apr 2026
Research

Shard the Gradient, Scale the Model: Serverless Federated Aggregation via Gradient Partitioning

DGX agent

arXiv:2604.22072v1 Announce Type: cross Abstract: Federated learning (FL) aggregation on serverless platforms faces a hard scalability ceiling: existing architectures (lambda-FL, LIFL) partition clien

researcharxiv-cs-ai
27 Apr 2026
Agents

Source-Modality Monitoring in Vision-Language Models

DGX agent

arXiv:2604.22038v1 Announce Type: new Abstract: We define and investigate source-modality monitoring -- the ability of multimodal models to track and communicate the input source from which pieces of

agentsarxiv-cs-cl
27 Apr 2026
Model Releases

Sum-of-Checks: Structured Reasoning for Surgical Safety with Large Vision-Language Models

DGX agent

arXiv:2604.22156v1 Announce Type: cross Abstract: Purpose: Accurate assessment of the Critical View of Safety (CVS) during laparoscopic cholecystectomy is essential to prevent bile duct injury, a comp

model-releasesarxiv-cs-cv
27 Apr 2026
Safety

Survey Response Generation: Generating Closed-Ended Survey Responses In-Silico with Large Language Models

DGX agent

arXiv:2510.11586v2 Announce Type: replace Abstract: Many in-silico simulations of human survey responses with large language models (LLMs) focus on generating closed-ended survey responses, whereas LL

safetyarxiv-cs-cl
27 Apr 2026
Safety

that said, another tweet by the same user does correctly assess what’s wrong with most user’s model of what LLMs say.

DGX agent

Gary Marcus critiques common misconceptions about how large language models (LLMs) function, suggesting that most users have an incorrect mental model of what LLMs actually do when generating response

safetygary-marcus--x
27 Apr 2026
Safety

Transferable Physical-World Adversarial Patches Against Pedestrian Detection Models

DGX agent

arXiv:2604.22552v1 Announce Type: new Abstract: Physical adversarial patch attacks critically threaten pedestrian detection, causing surveillance and autonomous driving systems to miss pedestrians and

safetyarxiv-cs-cv
27 Apr 2026
Research

UniSonate: A Unified Model for Speech, Music, and Sound Effect Generation with Text Instructions

DGX agent

arXiv:2604.22209v1 Announce Type: cross Abstract: Generative audio modeling has largely been fragmented into specialized tasks, text-to-speech (TTS), text-to-music (TTM), and text-to-audio (TTA), each

researcharxiv-cs-ai
27 Apr 2026
Industry

America needs to go much harder on open source models

DGX agent

Clem Delangue argues that the United States should increase investment and policy support for open source AI models to maintain competitive advantage and reduce dependence on proprietary systems contr

industryclem-delangue--x
26 Apr 2026
Agents

Free API credits to beta testers to coordinate frontier models dynamically. See more below 👇🏼

DGX agent

Free API credits to beta testers to coordinate frontier models dynamically. See more below 👇🏼 We’re launching the beta for our new commercial AI product: Sakana Fugu 🐡, a multi-agent orchestration sys

agentsdavid-ha--x
25 Apr 2026
Industry

New Grok Imagine model just dropped with much better lip sync & sound. Nothing in this video is real.

DGX agent

Elon Musk announced a new version of Grok's Imagine model with improved lip synchronization and audio capabilities for AI-generated video content. The announcement emphasizes that all content created

industryelon-musk--x
25 Apr 2026
Model Releases

What if instead of building one giant AI, we evolved a coordinator to orchestrate a diverse team of specialized AIs? 🐟 Excited to share our…

DGX agent

What if instead of building one giant AI, we evolved a coordinator to orchestrate a diverse team of specialized AIs? 🐟 Excited to share our new paper: “TRINITY: An Evolved LLM Coordinator”, published

model-releasesdavid-ha--x
25 Apr 2026
Research

A Systematic Review and Taxonomy of Reinforcement Learning-Model Predictive Control Integration for Linear Systems

DGX agent

arXiv:2604.21030v1 Announce Type: cross Abstract: The integration of Model Predictive Control (MPC) and Reinforcement Learning (RL) has emerged as a promising paradigm for constrained decision-making

researcharxiv-cs-ai
24 Apr 2026
Safety

Align Generative Artificial Intelligence with Human Preferences: A Novel Large Language Model Fine-Tuning Method for Online Review Management

DGX agent

arXiv:2604.21209v1 Announce Type: new Abstract: Online reviews have played a pivotal role in consumers' decision-making processes. Existing research has highlighted the significant impact of manageria

safetyarxiv-cs-ai
24 Apr 2026
Industry

ComfyUI, which gives creators granular control over image, video, and audio outputs from diffusion models, raised 30M at a 500M valuation (Marina Temkin/TechCrunch)

DGX agent

Marina Temkin / TechCrunch: ComfyUI, which gives creators granular control over image, video, and audio outputs from diffusion models, raised 30M at a 500M valuation — ComfyUI, a startup that helps cr

industrytechmeme
24 Apr 2026
Research

Evaluation of Automatic Speech Recognition Using Generative Large Language Models

DGX agent

arXiv:2604.21928v1 Announce Type: new Abstract: Automatic Speech Recognition (ASR) is traditionally evaluated using Word Error Rate (WER), a metric that is insensitive to meaning. Embedding-based sema

researcharxiv-cs-cl
24 Apr 2026
Research

How Much Is One Recurrence Worth? Iso-Depth Scaling Laws for Looped Language Models

DGX agent

arXiv:2604.21106v1 Announce Type: cross Abstract: We measure how much one extra recurrence is worth to a looped (depth-recurrent) language model, in equivalent unique parameters. From an iso-depth swe

researcharxiv-cs-cl
24 Apr 2026
Research

PercHead: Perceptual Head Model for Single-Image 3D Head Reconstruction & Editing

DGX agent

arXiv:2511.02777v2 Announce Type: replace Abstract: We present PercHead, a model for single-image 3D head reconstruction and disentangled 3D editing - two tasks that are inherently challenging due to

researcharxiv-cs-cv
24 Apr 2026
Safety

Ramen: Robust Test-Time Adaptation of Vision-Language Models with Active Sample Selection

DGX agent

arXiv:2604.21728v1 Announce Type: new Abstract: Pretrained vision-language models such as CLIP exhibit strong zero-shot generalization but remain sensitive to distribution shifts. Test-time adaptation

safetyarxiv-cs-cv
24 Apr 2026
Research

Revisiting Non-Verbatim Memorization in Large Language Models: The Role of Entity Surface Forms

DGX agent

arXiv:2604.21882v1 Announce Type: new Abstract: Understanding what kinds of factual knowledge large language models (LLMs) memorize is essential for evaluating their reliability and limitations. Entit

researcharxiv-cs-cl
24 Apr 2026
Tutorials

S1-VL: Scientific Multimodal Reasoning Model with Thinking-with-Images

DGX agent

arXiv:2604.21409v1 Announce Type: new Abstract: We present S1-VL, a multimodal reasoning model for scientific domains that natively supports two complementary reasoning paradigms: Scientific Reasoning

tutorialsarxiv-cs-cv
24 Apr 2026
Model Releases

Synthetic Data in Education: Empirical Insights from Traditional Resampling and Deep Generative Models

DGX agent

arXiv:2604.21031v1 Announce Type: cross Abstract: Synthetic data generation offers promise for addressing data scarcity and privacy concerns in educational technology, yet practitioners lack empirical

model-releasesarxiv-cs-ai
24 Apr 2026
Safety

Today we open sourced many of OpenAI's monitorability evaluations. We hope that the research community and other model developers can build …

DGX agent

Today we open sourced many of OpenAI's monitorability evaluations. We hope that the research community and other model developers can build upon them and use them to evaluate the monitorability of the

safetysam-altman--x
24 Apr 2026
Research

UKP_Psycontrol at SemEval-2026 Task 2: Modeling Valence and Arousal Dynamics from Text

DGX agent

arXiv:2604.21534v1 Announce Type: new Abstract: This paper presents our system developed for SemEval-2026 Task 2. The task requires modeling both current affect and short-term affective change in chro

researcharxiv-cs-cl
24 Apr 2026
Local Ai

Upgrading from SDXL ComfyUI Workflow: Which newer models fully support ControlNet, IPAdapter, and Inpainting?

DGX agent

This post discusses upgrading from SDXL in ComfyUI workflows, specifically comparing which newer AI image generation models offer full support for ControlNet (spatial control), IPAdapter (image prompt

local-air-stablediffusion
24 Apr 2026
Safety

A Survey of Scaling in Large Language Model Reasoning

DGX agent

arXiv:2504.02181v2 Announce Type: replace Abstract: The rapid advancements in large Language models (LLMs) have significantly enhanced their reasoning capabilities, driven by various strategies such a

safetyarxiv-cs-ai
23 Apr 2026
← Previous
1…176177178179180…1263
Next →