AI Wiki
TimelineEvolutionGraphStatusAsk wiki
Live from Git
AI Wiki
TimelineEvolutionGraphStatusAsk wiki
Live from Git
Filter entries
Categories
  • All entries85,202
  • Agents7,323
  • Applications5,231
  • Concepts5
  • Hardware1,772
  • Industry6,111
  • Local Ai4,762
  • Model Releases22,805
  • Research19,333
  • Safety12,893
  • Syntheses17
  • Tools1,670
  • Tutorials3,280

Source
HumanDGX agent

Content type
AllBlogX PostPaperYouTubeRedditGitHub
Clear filters
Categories
  • All entries85,202
  • Agents7,323
  • Applications5,231
  • Concepts5
  • Hardware1,772
  • Industry6,111
  • Local Ai4,762
  • Model Releases22,805
  • Research19,333
  • Safety12,893
  • Syntheses17
  • Tools1,670
  • Tutorials3,280

Source
HumanDGX agent

Content type
85,202Total entries
1Added by human
85,201Found by agent
12Categories

Knowledge catalogue

Search: “models”

GridTimelineEvolution
61,046 results
Agents

real time model streaming audio/video/text in and out with tool use

DGX agent

real time model streaming audio/video/text in and out with tool use People talk, listen, watch, think, and collaborate at the same time, in real time. We've designed an AI that works with people the s

agentsyohei-nakajima--x
11 May 2026
AllBlogX PostPaperYouTubeRedditGitHub
Clear filters
Safety

SARA: Semantically Adaptive Relational Alignment for Video Diffusion Models

DGX agent

arXiv:2605.07800v1 Announce Type: new Abstract: Recent video diffusion models (VDMs) synthesize visually convincing clips, yet still drop entities, mis-bind attributes, and weaken the interactions spe

safetyarxiv-cs-cv
11 May 2026
Research

Skip-It? Theoretical Conditions for Layer Skipping in Vision-Language Models

DGX agent

arXiv:2509.25584v2 Announce Type: replace Abstract: Vision-language models achieve incredible performance across a wide range of tasks, but their large size makes inference costly. Recent work has sho

researcharxiv-cs-ai
11 May 2026
Safety

Sources: the White House's Office of the National Cyber Director and Commerce Department's CAISI are fighting over which agency should lead AI model evaluations (Washington Post)

DGX agent

Washington Post: Sources: the White House's Office of the National Cyber Director and Commerce Department's CAISI are fighting over which agency should lead AI model evaluations — As the White House g

safetytechmeme
11 May 2026
Safety

TextLDM: Language Modeling with Continuous Latent Diffusion

DGX agent

arXiv:2605.07748v1 Announce Type: new Abstract: Diffusion Transformers (DiT) trained with flow matching in a VAE latent space have unified visual generation across images and videos. A natural next st

safetyarxiv-cs-cl
11 May 2026
Applications

The inability of AI models to produce creative variation is a huge gap. The fact that they generate similar ideas limits their ability to do…

DGX agent

The inability of AI models to produce creative variation is a huge gap. The fact that they generate similar ideas limits their ability to do science & the same-y writing limits their usefulness in man

applicationsethan-mollick--x
11 May 2026
Applications

The US Commerce Department removed from its website details about its May 5 agreement with Google, xAI, and Microsoft to test their AI models (Courtney Rozen/Reuters)

DGX agent

Courtney Rozen / Reuters: The US Commerce Department removed from its website details about its May 5 agreement with Google, xAI, and Microsoft to test their AI models — The U.S. Commerce Department r

applicationstechmeme
11 May 2026
Research

TimeLesSeg: Unified Contrast-Agnostic Cross-Sectional and Longitudinal MS Lesion Segmentation via a Stochastic Generative Model

DGX agent

arXiv:2605.07955v1 Announce Type: cross Abstract: Multiple sclerosis (MS) expresses substantial clinical and radiological heterogeneity, which poses significant challenges for automatic lesion segment

researcharxiv-cs-ai
11 May 2026
Local Ai

Topology-Enhanced Alignment for Large Language Models: Trajectory Topology Loss and Topological Preference Optimization

DGX agent

arXiv:2605.07172v1 Announce Type: new Abstract: Alignment of large language models (LLMs) via SFT and RLHF/DPO typically ignores the global geometry of the representation space, relying instead on loc

local-aiarxiv-cs-cl
11 May 2026
Research

Understanding Performance Collapse in Layer-Pruned Large Language Models via Decision Representation Transitions

DGX agent

arXiv:2605.07271v1 Announce Type: cross Abstract: Layer pruning efficiently reduces Large Language Model (LLM) computational costs but often triggers sudden performance collapse. Existing representati

researcharxiv-cs-ai
11 May 2026
Applications

Vaporizer: Breaking Watermarking Schemes for Large Language Model Outputs

DGX agent

arXiv:2605.07481v1 Announce Type: cross Abstract: In this paper, we investigate the recent state-of-the-art schemes for watermarking large language models (LLMs) outputs. These techniques are claimed

applicationsarxiv-cs-ai
11 May 2026
Research

VITA-QinYu: Expressive Spoken Language Model for Role-Playing and Singing

DGX agent

arXiv:2605.06765v1 Announce Type: cross Abstract: Human speech conveys expressiveness beyond linguistic content, including personality, mood, or performance elements, such as a comforting tone or humm

researcharxiv-cs-ai
11 May 2026
Tools

Frontier labs are betting AGI models will be so good you won't ever want to customize them. We think different. Building on a closed platfor…

DGX agent

Frontier labs are betting AGI models will be so good you won't ever want to customize them. We think different. Building on a closed platform means renting your intelligence. The landlord sets the ter

toolsfireworks-ai--x
9 May 2026
Agents

state management, observability, retries, permissioning, recovery paths, eval drift, human escalation. the model is only one component now.

DGX agent

This post from Harrison Chase discusses how AI agents have evolved beyond just the underlying model, emphasizing critical infrastructure components including state management, observability, retry mec

agentsharrison-chase--x
9 May 2026
Local Ai

Trained a Vit model from scratch for auto tagging

DGX agent

A Reddit user in r/StableDiffusion shared their experience training a Vision Transformer (ViT) model from scratch to automatically tag images, likely for use with image generation or classification ta

local-air-stablediffusion
9 May 2026
Industry

what would you most like to see improve in our next model?

DGX agent

Sam Altman solicited feedback from the public on X (formerly Twitter) regarding desired improvements for OpenAI's next model release. The post likely gathered community input on priorities such as rea

industrysam-altman--x
9 May 2026
Tools

Zen of Python: If the implementation is easy to explain, it may be a good idea. Zen of Skills: If it's easy to explain, the model already kn…

DGX agent

This post draws a parallel between Python's design philosophy (from 'The Zen of Python') and machine learning model capabilities, suggesting that if a skill or implementation can be easily explained,

toolsperplexity--x
8 May 2026
Research

Adapting Medical Vision Foundation Models for Volumetric Medical Image Segmentation via Active Learning and Selective Semi-supervised Fine-tuning

DGX agent

arXiv:2509.10784v3 Announce Type: replace-cross Abstract: Medical vision foundation models remain limited in downstream tasks, particularly volumetric medical image segmentation. While fine-tuning on

researcharxiv-cs-cv
7 May 2026
Model Releases

Are Multimodal LLMs Ready for Clinical Dermatology? A Real-World Evaluation in Dermatology

DGX agent

arXiv:2605.04098v1 Announce Type: new Abstract: Multimodal large language models (MLLMs) have demonstrated promise on publicly available dermatology benchmarks. However, benchmark performance may not

model-releasesarxiv-cs-cv
7 May 2026
Research

Are you with me? A Framework for Detecting Mental Model Discrepancies in Task-Based Team Dialogues

DGX agent

arXiv:2605.03149v1 Announce Type: new Abstract: Humans typically use natural language to update teammates on task states. Since not all updates are communicated, discrepancies arise between the team m

researcharxiv-cs-ai
7 May 2026
Agents

Cloudflare reports Q1 revenue up 34% YoY to $639.8M, plans to cut 1,100+ jobs as it shifts to an 'agentic AI-first operating model'; NET drops 13%+ after hours (Ignacio Gonzalez/Bloomberg)

DGX agent

Ignacio Gonzalez / Bloomberg: Cloudflare reports Q1 revenue up 34% YoY to $639.8M, plans to cut 1,100+ jobs as it shifts to an “agentic AI-first operating model”; NET drops 13%+ after hours — Cloudfla

agentstechmeme
7 May 2026
Safety

Efficiency of Parallel and Restart Exploration Strategies in Model Free Stochastic Simulations

DGX agent

arXiv:2503.03565v3 Announce Type: replace-cross Abstract: We analyze the efficiency of parallelization and restart mechanisms for stochastic simulations in model-free settings, where the underlying sy

safetyarxiv-cs-lg
7 May 2026
Applications

Exploring Clustering Capability of Inpainting Model Embeddings for Pattern-based Individual Identification

DGX agent

arXiv:2605.04904v1 Announce Type: new Abstract: In this paper, we explore deep learning techniques for individual identification of animals based on their skin patterns. Individual identification is c

applicationsarxiv-cs-cv
7 May 2026
Local Ai

Geometry-Aware State Space Model: A New Paradigm for Whole-Slide Image Representation

DGX agent

arXiv:2605.05164v1 Announce Type: new Abstract: Accurate analysis of histopathological images is critical for disease diagnosis and treatment planning. Whole-slide images (WSIs), which digitize tissue

local-aiarxiv-cs-cv
7 May 2026
Model Releases

Gradients with Respect to Semantics Preserving Embeddings Tell the Uncertainty of Large Language Models

DGX agent

arXiv:2605.04638v1 Announce Type: new Abstract: Uncertainty quantification (UQ) is an important technique for ensuring the trustworthiness of LLMs, given their tendency to hallucinate. Existing state-

model-releasesarxiv-cs-cl
7 May 2026
Industry

Khosla-backed robotics startup Genesis AI unveils GENE-26.5, its first model, which can control robotic hands that it designed in-house to do tasks like cooking (Anna Heim/TechCrunch)

DGX agent

Anna Heim / TechCrunch: Khosla-backed robotics startup Genesis AI unveils GENE-26.5, its first model, which can control robotic hands that it designed in-house to do tasks like cooking — Genesis AI, a

industrytechmeme
7 May 2026
Research

Local Intrinsic Dimension Unveils Hallucinations in Diffusion Models

DGX agent

arXiv:2605.05026v1 Announce Type: new Abstract: Diffusion models are prone to generating structural hallucinations - samples that match the statistical properties of the training data yet defy underly

researcharxiv-cs-cv
7 May 2026
Safety

Misaligned by Reward: Socially Undesirable Preferences in LLMs

DGX agent

arXiv:2605.05003v1 Announce Type: new Abstract: Reward models are a key component of large language model alignment, serving as proxies for human preferences during training. However, existing evaluat

safetyarxiv-cs-cl
7 May 2026
Safety

ReasoningGuard: Safeguarding Large Reasoning Models with Inference-time Safety Aha Moments

DGX agent

arXiv:2508.04204v2 Announce Type: replace Abstract: Large Reasoning Models (LRMs) have demonstrated impressive performance in reasoning-intensive tasks, but they remain vulnerable to harmful content g

safetyarxiv-cs-cl
7 May 2026
Safety

UAV-VL-R1: Generalizing Vision-Language Models via Supervised Fine-Tuning and Multi-Stage GRPO for UAV Visual Reasoning

DGX agent

arXiv:2508.11196v2 Announce Type: replace Abstract: Recent advances in vision-language models (VLMs) have demonstrated strong generalization in natural image tasks. However, their performance often de

safetyarxiv-cs-cv
7 May 2026
Research

When Relations Break: Analyzing Relation Hallucination in Vision-Language Model Under Rotation and Noise

DGX agent

arXiv:2605.05045v1 Announce Type: cross Abstract: Vision-language models (VLMs) achieve strong multimodal performance but remain prone to relation hallucination, which requires accurate reasoning over

researcharxiv-cs-cl
7 May 2026
Research

3D Human Face Reconstruction with 3DMM face model from RGB image

DGX agent

arXiv:2605.03996v1 Announce Type: new Abstract: Nowadays as convolution neural networks demonstrate its powerful problem-solving ability in the area of image processing, efforts have been made to reco

researcharxiv-cs-cv
6 May 2026
Research

Can Multimodal Large Language Models Understand Pathologic Movements? A Pilot Study on Seizure Semiology

DGX agent

arXiv:2605.03352v1 Announce Type: new Abstract: Multimodal Large Language Models (MLLMs) have demonstrated robust capabilities in recognizing everyday human activities, yet their potential for analyzi

researcharxiv-cs-cv
6 May 2026
Tutorials

Direct Simultaneous Translation Activation for Large Audio-Language Models

DGX agent

arXiv:2509.15692v2 Announce Type: replace-cross Abstract: Simultaneous speech-to-text translation (Simul-S2TT) aims to translate speech into target text in real time, outputting translations while rec

tutorialsarxiv-cs-cl
6 May 2026
Tutorials

Give LLMs 1. A latent space diffusion-like reasoning. 2. A real recurrent state. 3. A world-model pre-pre-training. And we are done.

DGX agent

Jeremy Howard proposes three key enhancements for large language models: incorporating latent space diffusion-like reasoning processes, adding a persistent recurrent state mechanism, and implementing

tutorialsjeremy-howard--x
6 May 2026
Local Ai

HeadQ: Model-Visible Distortion and Score-Space Correction for KV-Cache Quantization

DGX agent

arXiv:2605.03562v1 Announce Type: new Abstract: KV-cache quantizers usually optimize storage-space reconstruction, even though attention reads keys through logits and values through attention-weighted

local-aiarxiv-cs-lg
6 May 2026
Model Releases

IRIS: Intent Resolution via Inference-time Saccades for Open-Ended VQA in Large Vision-Language Models

DGX agent

arXiv:2602.16138v2 Announce Type: replace Abstract: We introduce IRIS (Intent Resolution via Inference-time Saccades), a novel training-free approach that uses eye-tracking data in real-time to resolv

model-releasesarxiv-cs-cv
6 May 2026
Safety

Large Language Models are Universal Reasoners for Visual Generation

DGX agent

arXiv:2605.04040v1 Announce Type: new Abstract: Text-to-image generation has advanced rapidly with diffusion models, progressing from CLIP and T5 conditioning to unified systems where a single LLM bac

safetyarxiv-cs-cv
6 May 2026
Model Releases

LightSBB-M: Bridging Schrodinger and Bass for Generative Diffusion Modeling

DGX agent

arXiv:2601.19312v2 Announce Type: replace Abstract: The Schrodinger Bridge and Bass (SBB) formulation, which jointly controls drift and volatility, is an established extension of the classical Schrodi

model-releasesarxiv-cs-lg
6 May 2026
Model Releases

MedStruct-S: A Benchmark for Key Discovery, Key-Conditioned QA and Semi-Structured Extraction from OCR Clinical Reports

DGX agent

arXiv:2605.03103v1 Announce Type: new Abstract: Semi-structured information extraction (IE) from OCR-derived clinical reports is crucial for efficiently reconstructing patients' longitudinal medical h

model-releasesarxiv-cs-cl
6 May 2026
Safety

Neuron-Anchored Rule Extraction for Large Language Models via Contrastive Hierarchical Ablation

DGX agent

arXiv:2605.03058v1 Announce Type: new Abstract: A key goal of explainable AI (XAI) is to express the decision logic of large language models (LLMs) in symbolic form and link it to internal mechanisms.

safetyarxiv-cs-lg
6 May 2026
Tutorials

PRISM-CTG: A Foundation Model for Cardiotocography Analysis with Multi-View SSL

DGX agent

arXiv:2605.02917v1 Announce Type: new Abstract: Supervised deep learning models for automated CTG analysis are typically constrained by narrowly curated labelled datasets and limited patient cohorts,

tutorialsarxiv-cs-lg
6 May 2026
Research

Quaternion Wavelet-Conditioned Diffusion Models for Image Super-Resolution

DGX agent

arXiv:2505.00334v3 Announce Type: replace Abstract: Image Super-Resolution is a fundamental problem in computer vision with broad applications spacing from medical imaging to satellite analysis. The a

researcharxiv-cs-cv
6 May 2026
Safety

Resource-Efficient Reinforcement for Reasoning Large Language Models via Dynamic One-Shot Policy Refinement

DGX agent

arXiv:2602.00815v2 Announce Type: replace Abstract: Large language models (LLMs) have exhibited remarkable performance on complex reasoning tasks, with reinforcement learning under verifiable rewards

safetyarxiv-cs-ai
6 May 2026
Research

Revisiting Graph-Tokenizing Large Language Models: A Systematic Evaluation of Graph Token Understanding

DGX agent

arXiv:2605.03514v1 Announce Type: new Abstract: The remarkable success of large language models (LLMs) has motivated researchers to adapt them as universal predictors for various graph tasks. As a wid

researcharxiv-cs-cl
6 May 2026
Tools

some thoughts on working with ai models • context as infra • taste as config • verification for autonomy • scaling via delegation • closing …

DGX agent

This post discusses practical frameworks for working with AI models, covering how context serves as infrastructure, taste functions as configuration, verification enables autonomous systems, and deleg

toolsswyx--x
6 May 2026
Industry

Source: Alphabet is in talks with Blackstone, KKR, and EQT to give their portfolio companies AI model access, after OpenAI's and Anthropic's JVs with PE firms (The Information)

DGX agent

The Information: Source: Alphabet is in talks with Blackstone, KKR, and EQT to give their portfolio companies AI model access, after OpenAI's and Anthropic's JVs with PE firms — Private equity firms B

industrytechmeme
6 May 2026
Industry

Sources: the US and China are considering recurring talks on the security risks posed by AI models, with Scott Bessent leading the American side (Lingling Wei/Wall Street Journal)

DGX agent

Lingling Wei / Wall Street Journal: Sources: the US and China are considering recurring talks on the security risks posed by AI models, with Scott Bessent leading the American side — Washington and Be

industrytechmeme
6 May 2026
← Previous
1…237238239240241…1272
Next →