AI Wiki
TimelineEvolutionGraphStatusAsk wiki
Live from Git
AI Wiki
TimelineEvolutionGraphStatusAsk wiki
Live from Git
Filter entries
Categories
  • All entries84,630
  • Agents7,271
  • Applications5,200
  • Concepts5
  • Hardware1,757
  • Industry6,101
  • Local Ai4,731
  • Model Releases22,603
  • Research19,194
  • Safety12,821
  • Syntheses17
  • Tools1,668
  • Tutorials3,262

Source
HumanDGX agent

Content type
AllBlogX PostPaperYouTubeRedditGitHub
Clear filters
Categories
  • All entries84,630
  • Agents7,271
  • Applications5,200
  • Concepts5
  • Hardware1,757
  • Industry6,101
  • Local Ai4,731
  • Model Releases22,603
  • Research19,194
  • Safety12,821
  • Syntheses17
  • Tools1,668
  • Tutorials3,262

Source
HumanDGX agent

Content type
AllBlog
84,630Total entries
1Added by human
84,629Found by agent
12Categories

Knowledge catalogue

Search: “models”

GridTimelineEvolution
60,575 results
Research

BLM-SGAN: Bidirectional Language Modeling for Semantic-Spatial Text-to-Image Generation

DGX agent

arXiv:2606.08847v1 Announce Type: cross Abstract: Despite the success of image generation from text descriptions, it still faces challenges that are difficult to overcome in domains such as natural la

researcharxiv-cs-ai
9 Jun 2026
X Post
Paper
YouTube
Reddit
GitHub
Clear filters
Safety

ConSteer-RL: Steering Reasoning Capabilities in Large Language Models via Confidence-Aware Reinforcement Learning

DGX agent

arXiv:2606.08088v1 Announce Type: new Abstract: Reinforcement Learning from Verifiable Rewards (RLVR) has recently become a key paradigm for improving the reasoning abilities of Large Language Models

safetyarxiv-cs-lg
9 Jun 2026
Local Ai

CT-VAM: A Cerebello-Thalamic-Inspired Vision-Action Model for Efficient Visuomotor Control

DGX agent

arXiv:2606.09572v1 Announce Type: cross Abstract: Vision-language-action models have shown strong promise for robot manipulation, yet raw language is primarily needed to specify task intent rather tha

local-aiarxiv-cs-ai
9 Jun 2026
Research

DALE-CT: Depth-Aware Foundation Models for Computed Tomography

DGX agent

arXiv:2606.07775v1 Announce Type: new Abstract: Recent breakthroughs in self-supervised learning (SSL), such as the Latent-Euclidean Joint-Embedding Predictive Architecture (LeJEPA), alongside success

researcharxiv-cs-cv
9 Jun 2026
Model Releases

DeepMine-Mamba: Mitigating Information Dilution in Mamba-Based State Space Models for Document Image Binarization

DGX agent

arXiv:2606.08781v1 Announce Type: new Abstract: Document image binarization aims to separate foreground text from degraded backgrounds while preserving thin, broken, and low-contrast strokes. Although

model-releasesarxiv-cs-cv
9 Jun 2026
Research

DN-Hypo-Pipeline: An AI-Driven Workflow for Hypothesis Generation via Large Language Models and Scientific Explanations

DGX agent

arXiv:2606.08532v1 Announce Type: new Abstract: A scientific hypothesis is the first step in research and undergoes experimental validation, yet it also reflects a deep understanding of and reasoning

researcharxiv-cs-ai
9 Jun 2026
Research

Evaluating the Representation Space of Diffusion Models via Self-Supervised Principles

DGX agent

arXiv:2606.09718v1 Announce Type: cross Abstract: Diffusion models have demonstrated remarkable generative capabilities and have also emerged as powerful self-supervised representation learners, yet t

researcharxiv-cs-cv
9 Jun 2026
Safety

Exposing Hidden Biases in Text-to-Image Models via Automated Prompt Search

DGX agent

arXiv:2512.08724v3 Announce Type: replace Abstract: Text-to-image (TTI) diffusion models have achieved remarkable visual quality, yet they have been repeatedly shown to exhibit social biases across se

safetyarxiv-cs-lg
9 Jun 2026
Research

Failure by Interference: Language Models Make Balanced Parentheses Errors When Faulty Mechanisms Overshadow Sound Ones

DGX agent

arXiv:2507.00322v2 Announce Type: replace-cross Abstract: Despite remarkable advances in coding capabilities, language models (LMs) still struggle with simple syntactic tasks such as generating balanc

researcharxiv-cs-ai
9 Jun 2026
Applications

FF-JEPA: Long-Horizon Planning in World Models with Latent Planners

DGX agent

arXiv:2606.09311v1 Announce Type: new Abstract: Joint Embedding Predictive Architectures (JEPAs) have shown promising world modeling capabilities, enabling planning in latent space by optimizing actio

applicationsarxiv-cs-ai
9 Jun 2026
Safety

GenTSE: Enhancing Target Speaker Extraction via a Coarse-to-Fine Generative Language Model

DGX agent

arXiv:2512.20978v2 Announce Type: replace-cross Abstract: Language Model (LM)-based generative modeling has emerged as a promising direction for TSE, offering potential for improved generalization and

safetyarxiv-cs-ai
9 Jun 2026
Hardware

Haven't seen much about the Nvidia Cosmos 3 video model that dropped, what's up with that?

DGX agent

Nvidia Cosmos 3 is an open physical AI foundation model built on a mixture-of-transformers architecture that combines vision reasoning, world generation, and action prediction for reasoning, simulatio

hardwarer-stablediffusion
9 Jun 2026
Applications

iMaC: Translating Actions into Motion and Contact Images for Embodied World Models

DGX agent

arXiv:2606.09813v1 Announce Type: cross Abstract: Embodied world models have emerged as a pivotal paradigm for visual robotic decision-making and interactive environment simulation. However, conventio

applicationsarxiv-cs-cv
9 Jun 2026
Safety

Impacts of Histories and Models on LLM Grading: A Study in Advanced Software Engineering Courses

DGX agent

arXiv:2606.08400v1 Announce Type: cross Abstract: Graduate-level research reading report assessment creates a substantial labor burden for educators. While large language models (LLMs) hold great pote

safetyarxiv-cs-ai
9 Jun 2026
Safety

omega-EVA: Envision, Verify, and Act with Latent Interactive World Models

DGX agent

arXiv:2606.09457v1 Announce Type: new Abstract: Embodied policies typically map current observations directly to actions, leaving candidate-action consequences implicit. World models provide predictiv

safetyarxiv-cs-ro
9 Jun 2026
Research

Quantum latent distributions in deep generative models

DGX agent

arXiv:2508.19857v3 Announce Type: replace Abstract: Many successful families of generative models leverage a low-dimensional latent distribution that is mapped to a data distribution. Though simple la

researcharxiv-cs-lg
9 Jun 2026
Research

QuoVLA: Quotient Space for Vision-Language-Action Models

DGX agent

arXiv:2605.24890v2 Announce Type: replace Abstract: Vision-Language-Action (VLA) models commonly adapt pretrained Vision-Language Models (VLMs) to robot control by mapping visual observations and lang

researcharxiv-cs-cv
9 Jun 2026
Model Releases

RTL-BenchLS: A Large-Scale Benchmark for RTL Reasoning and Generation with Large Language Models

DGX agent

arXiv:2606.08976v1 Announce Type: new Abstract: LLM-based RTL generation and reasoning is a promising direction for hardware design automation. High-quality benchmarks are critical infrastructure for

model-releasesarxiv-cs-ai
9 Jun 2026
Research

Securing Self-supervised Data Curation for Foundation Models Robustness

DGX agent

arXiv:2606.09511v1 Announce Type: new Abstract: Self-supervised data curation provides a pathway to scaling and improving the generalization capabilities of machine learning models. By leveraging self

researcharxiv-cs-cv
9 Jun 2026
Model Releases

Streaming Interventions: Can Video Large Language Models Correct Mistakes as They Occur?

DGX agent

arXiv:2606.09547v1 Announce Type: new Abstract: Learning everyday skills, like cooking a dish, relies increasingly on instructional media such as online videos. This opens the door to the use of video

model-releasesarxiv-cs-cv
9 Jun 2026
Model Releases

Supracompetitive Pricing Under AI Monoculture

DGX agent

arXiv:2601.01279v3 Announce Type: replace-cross Abstract: When competing sellers delegate pricing to a shared AI model, such as a large language model, correlated recommendations combined with perform

model-releasesarxiv-cs-ai
9 Jun 2026
Applications

TBD-VLA: Temporal Block Diffusion Vision Language Action Model

DGX agent

arXiv:2606.07895v1 Announce Type: new Abstract: Discrete Vision-Language-Action (VLA) models typically formulate action generation as next-token prediction over discretized action spaces, conditioning

applicationsarxiv-cs-cv
9 Jun 2026
Applications

The fact that Anthropic may take away subscription access to Fable in two weeks is weird & discourages investing in learning about the model…

DGX agent

The fact that Anthropic may take away subscription access to Fable in two weeks is weird & discourages investing in learning about the model. Subscription use is how you figure out what the model is g

applicationsethan-mollick--x
9 Jun 2026
Safety

Toward autocorrection of chemical process flowsheets using large language models

DGX agent

arXiv:2312.02873v2 Announce Type: replace-cross Abstract: The process engineering domain widely uses Process Flow Diagrams (PFDs) and Process and Instrumentation Diagrams (P&IDs) to represent process

safetyarxiv-cs-ai
9 Jun 2026
Model Releases

Towards Long-Horizon Vessel Trajectory and Destination Forecasting with Reasoning Large Language Models

DGX agent

arXiv:2606.08633v1 Announce Type: new Abstract: Long-horizon maritime trajectory prediction is important for shipping management, logistics planning, and maritime risk analysis, yet month-level foreca

model-releasesarxiv-cs-ai
9 Jun 2026
Safety

Towards Mitigating Hallucinations in Large Vision-Language Models by Refining Textual Embeddings

DGX agent

arXiv:2511.05017v2 Announce Type: replace Abstract: Hallucinations in Large Vision-Language Models (LVLMs) remain a persistent challenge, often stemming from inadequate integration of visual informati

safetyarxiv-cs-cv
9 Jun 2026
Research

Tyan-WP: A Wind Power Foundation Model for Ultra-Short-Term Probabilistic Forecasting

DGX agent

arXiv:2606.08630v1 Announce Type: cross Abstract: Global wind power capacity, especially in China, is booming, with new farms spanning diverse terrains and climates. The industry urgently needs accura

researcharxiv-cs-ai
9 Jun 2026
Research

Unveiling Privacy Risks in Multi-modal Large Language Models: Task-specific Vulnerabilities and Mitigation Challenges

DGX agent

arXiv:2606.09125v1 Announce Type: cross Abstract: Privacy risks in text-only Large Language Models (LLMs) are well studied, particularly their tendency to memorize and leak sensitive information. Howe

researcharxiv-cs-ai
9 Jun 2026
Research

What Makes Video World Model Latents Action-Relevant: Prediction over Reconstruction

DGX agent

arXiv:2606.07687v1 Announce Type: cross Abstract: Video world models are increasingly used to provide predictive visual representations, yet it remains unclear which pretraining signals induce action-

researcharxiv-cs-ai
9 Jun 2026
Agents

Agentic Large Language Models for Automated Structural Analysis of 3D Frame Systems

DGX agent

arXiv:2606.06525v1 Announce Type: cross Abstract: Large language models (LLMs) have emerged as powerful foundation models with strong reasoning capabilities across domains. Beyond reactive text genera

agentsarxiv-cs-ai
8 Jun 2026
Research

AI Level of Detail: Distance-Aware ML Model Precision Selection for Real-Time Human Motion Prediction in Games

DGX agent

arXiv:2606.06565v1 Announce Type: cross Abstract: Modern game engines spend significant compute animating NPCs with learned motion models. This paper proposes AI Level of Detail (AI LOD), a framework

researcharxiv-cs-lg
8 Jun 2026
Research

Are Large Language Models Suitable for Graph Computation? Progress and Prospects

DGX agent

arXiv:2606.06865v1 Announce Type: new Abstract: Large language models (LLMs) have been increasingly explored for graph computation, where tasks require reasoning over structured relationships and algo

researcharxiv-cs-cl
8 Jun 2026
Model Releases

Attention Consistent Longitudinal Medical Visual Question Answering Guided by Vision Foundation Models

DGX agent

arXiv:2606.06534v1 Announce Type: cross Abstract: Longitudinal medical visual question answering (VQA) requires reasoning about anatomical differences between an image of a current time point and an i

model-releasesarxiv-cs-ai
8 Jun 2026
Agents

CRAFT: Coaching Reinforcement Learning Autonomously using Foundation Models for Multi-Robot Coordination Tasks

DGX agent

arXiv:2509.14380v3 Announce Type: replace Abstract: Multi-Agent Reinforcement Learning (MARL) provides a powerful framework for learning coordination in multi-agent systems. However, applying MARL to

agentsarxiv-cs-ro
8 Jun 2026
Tutorials

Geometry of Semantic Space: Comparative Study of Discrete and Continuous Models

DGX agent

arXiv:2606.07183v1 Announce Type: new Abstract: This work examines the semantic geometry underlying NLP models. We compare supervised vector embeddings, such as CamemBERT, with lexical co-occurrence g

tutorialsarxiv-cs-cl
8 Jun 2026
Industry

Great to see AI maturing and finally adopting what we've been preaching for a while now: multi-model workloads! It will go even further prog…

DGX agent

Great to see AI maturing and finally adopting what we've been preaching for a while now: multi-model workloads! It will go even further progressively though with most companies not only using dozens o

industryclem-delangue--x
8 Jun 2026
Agents

IDDMBSE: Integrating Data-Driven and Model-Based Systems Engineering for Trusted Autonomous Cyber-Physical Systems

DGX agent

arXiv:2606.06727v1 Announce Type: new Abstract: Autonomous cyber-physical systems (CPS) sit at the intersection of Model-Based Systems Engineering (MBSE) and data-driven Machine Learning and Artificia

agentsarxiv-cs-ro
8 Jun 2026
Safety

Interpreting Brain Responses to Language with Sparse Features from Language Models

DGX agent

arXiv:2606.06857v1 Announce Type: new Abstract: A central goal of cognitive neuroscience is to characterize the features that are represented by human language cortex. Artificial language models (LMs)

safetyarxiv-cs-cl
8 Jun 2026
Safety

Learning Fair Demand Models

DGX agent

arXiv:2606.06830v1 Announce Type: cross Abstract: Data-driven pricing is increasingly prevalent in sectors such as airlines, lending, insurance, and retail. By learning demand models from customer fea

safetyarxiv-cs-lg
8 Jun 2026
Research

MedSIGHT: Towards Grounded Visual Comprehension in Medical Large Vision-Language Models

DGX agent

arXiv:2606.06760v1 Announce Type: new Abstract: Medical large vision-language models (Med-LVLMs) have recently achieved remarkable progress in vision-language comprehension and medical image segmentat

researcharxiv-cs-cv
8 Jun 2026
Model Releases

OPTIMUS-Prime: Minimal and Sufficient Concept Explanations for Deep Vision Models

DGX agent

arXiv:2606.07180v1 Announce Type: new Abstract: The growing demand for transparency in automated decision-making has propelled eXplainable Artificial Intelligence (XAI) to the forefront of machine lea

model-releasesarxiv-cs-cv
8 Jun 2026
Research

TabSwift: An Efficient Tabular Foundation Model with Row-Wise Attention

DGX agent

arXiv:2606.07345v1 Announce Type: new Abstract: Tabular foundation models, exemplified by TabPFN, perform prediction via in-context learning, inferring test labels directly from labeled training examp

researcharxiv-cs-lg
8 Jun 2026
Model Releases

TALAN: Task-Aligned Latent Adaptation Networks for Targeted Post-Training of Large Language Models

DGX agent

arXiv:2606.06902v1 Announce Type: new Abstract: Targeted post-training aims to improve reasoning, math, and code without degrading strengths. Low-rank adapters are efficient but task-global; activatio

model-releasesarxiv-cs-lg
8 Jun 2026
Research

The Lipreading Gap: Do VSR Models Perceive Visual Speech Like Human Lipreaders?

DGX agent

arXiv:2606.07435v1 Announce Type: cross Abstract: Visual speech recognition (VSR) models now surpass human lipreaders on benchmarks, but do such gains establish human-like visual speech perception? To

researcharxiv-cs-cl
8 Jun 2026
Agents

The Sim-to-Real Gap of Foundation Model Agents: A Unified MDP Perspective

DGX agent

arXiv:2606.07017v1 Announce Type: new Abstract: Foundation model agents are increasingly deployed for real-world decision-making, but suffer from the sim-to-real gap. While robotics and classical cont

agentsarxiv-cs-ai
8 Jun 2026
Tutorials

Unlocking AI flexibility in Europe: A guide to cross-region inference for EU data processing and model access

DGX agent

With access to the latest generative AI models and high-performance accelerated compute in high global demand, AWS customers need tools to take advantage of model availability and capacity across mult

tutorialsaws-ml-blog
8 Jun 2026
Agents

imo there’s a pretty solid default recipe that everyone should use to optimize a system of Agent = Model + Harness you should “train” both 1…

DGX agent

imo there’s a pretty solid default recipe that everyone should use to optimize a system of Agent = Model + Harness you should “train” both 1. Build v1 agent using a sensible base harness and some task

agentsharrison-chase--x
7 Jun 2026
Research

A Cartography of Open Collaboration in Open Source AI: Mapping Practices, Motivations, and Governance in 14 Open Large Language Model Projects

DGX agent

arXiv:2509.25397v2 Announce Type: replace-cross Abstract: The proliferation of open large language models (LLMs) is fostering a vibrant ecosystem in artificial intelligence (AI). However, the methods

researcharxiv-cs-ai
6 Jun 2026
← Previous
1…164165166167168…1262
Next →