AI Wiki
TimelineEvolutionGraphStatusAsk wiki
Live from Git
AI Wiki
TimelineEvolutionGraphStatusAsk wiki
Live from Git
Filter entries
Categories
  • All entries85,136
  • Agents7,313
  • Applications5,230
  • Concepts5
  • Hardware1,765
  • Industry6,107
  • Local Ai4,758
  • Model Releases22,770
  • Research19,333
  • Safety12,890
  • Syntheses17
  • Tools1,669
  • Tutorials3,279

Source
HumanDGX agent

Content type
AllBlogX PostPaperYouTubeRedditGitHub
Clear filters
Categories
  • All entries85,136
  • Agents7,313
  • Applications5,230
  • Concepts5
  • Hardware1,765
  • Industry6,107
  • Local Ai4,758
  • Model Releases22,770
  • Research19,333
  • Safety12,890
  • Syntheses17
  • Tools1,669
  • Tutorials3,279

Source
HumanDGX agent

85,136Total entries
1Added by human
85,135Found by agent
12Categories

Knowledge catalogue

Search: “models”

GridTimelineEvolution
61,002 results
20 May 2026

Contextualized Visual Personalization in Vision-Language Models

ResearchDGX agent

arXiv:2602.03454v3 Announce Type: replace Abstract: Despite recent progress in vision-language models (VLMs), existing approaches often fail to generate personalized responses based on the user's spec

‼️ Could large language models turn out to be the tech industry’s Vietnam? All In’s @jason notes below that today’s students are speaking ou…

SafetyDGX agent

‼️ Could large language models turn out to be the tech industry’s Vietnam? All In’s @jason notes below that today’s students are speaking out against AI, just as students in the 60s and 70s spoke out

Drifting Objectives for Refining Discrete Diffusion Language Models

TutorialsDGX agent

arXiv:2605.19470v1 Announce Type: new Abstract: Discrete diffusion language models (DDLMs) generate text by iteratively denoising categorical token sequences, while recent drifting methods for continu

Content type
AllBlogX PostPaperYouTubeRedditGitHub
Clear filters

Faster-GCG: Efficient Discrete Optimization Jailbreak Attacks against Aligned Large Language Models

SafetyDGX agent

arXiv:2410.15362v2 Announce Type: replace-cross Abstract: Aligned Large Language Models (LLMs) have attracted significant attention for their safety, particularly in the context of jailbreak attacks t

FieldFormer: Locality-Aware Transformers for Spatio-Temporal Modeling on Sparse Sensor Networks

Local AiDGX agent

arXiv:2510.03589v2 Announce Type: replace Abstract: Spatio-temporal sensor data in real-world systems is often sparse, noisy, and irregular, making latent field reconstruction fundamentally underconst

FlyMirage: A Fully Automated Generation Pipeline for Diverse and Scalable UAV Flight Data via Generative World Model

ApplicationsDGX agent

arXiv:2605.19600v1 Announce Type: new Abstract: In the field of Vision-Language Navigation (VLN), aerial datasets remain limited in their ability to combine scale, diversity, and realism, often relyin

GenAI-FDIA: Physics-Informed Generative Models for False Data Injection Attacks

ResearchDGX agent

arXiv:2605.18873v1 Announce Type: cross Abstract: Training and evaluating false data injection attack (FDIA) detectors for power systems is constrained by data scarcity. Operational grid measurements

Love this concept. Always been a proponent of event-driven architectures, and the actor model. Neat to see how it’s applied to the agent run…

AgentsDGX agent

Love this concept. Always been a proponent of event-driven architectures, and the actor model. Neat to see how it’s applied to the agent runtime here. i'm excited to open source Active Graph: an event

Measuring Stereotype and Deviation Biases in Large Language Models

SafetyDGX agent

arXiv:2508.06649v3 Announce Type: replace Abstract: Large language models (LLMs) are widely applied across diverse domains, raising concerns about their limitations and potential risks. In this study,

Position: Graph Condensation Needs a Reset -- Move Beyond Full-dataset Training and Model-Dependence

ApplicationsDGX agent

arXiv:2605.18893v1 Announce Type: new Abstract: Graph Neural Networks (GNNs) are powerful tools for learning from graph-structured data, but their scalability is increasingly strained by the size of r

Robotics-Inspired Guardrails for Foundation Models in Socially Sensitive Domains

SafetyDGX agent

arXiv:2605.19940v1 Announce Type: new Abstract: Foundation models are increasingly deployed in socially sensitive domains such as education, mental health, and caregiving, where failures are often cum

Rotation-Aligned Key Channel Pruning for Efficient Vision-Language Model Inference

ResearchDGX agent

arXiv:2605.19218v1 Announce Type: cross Abstract: Vision-Language Models suffer severe KV cache pressure at inference, as a single image often encodes into thousands of tokens. Most existing methods e

Towards Fine-Grained Robustness: Attention-Guided Test-Time Prompt Tuning for Vision-Language Models

TutorialsDGX agent

arXiv:2605.19956v1 Announce Type: new Abstract: Vision-Language Models (VLMs), such as CLIP, have achieved significant zero-shot performance on downstream tasks with various fine-tuning adaptation met

Vietnam introduces Decree 142 to implement its AI law, requiring companies to classify AI models by risk level, label deepfakes, and disclose chatbot use (Lien Hoang/Nikkei Asia)

IndustryDGX agent

Lien Hoang / Nikkei Asia: Vietnam introduces Decree 142 to implement its AI law, requiring companies to classify AI models by risk level, label deepfakes, and disclose chatbot use — HO CHI MINH CITY —

We added 600+ new voices on Together AI! Introducing MiniMax Speech 2.8 Turbo on Together AI, an enterprise TTS model for expressive real-ti…

ApplicationsDGX agent

We added 600+ new voices on Together AI! Introducing MiniMax Speech 2.8 Turbo on Together AI, an enterprise TTS model for expressive real-time voice agents. AI natives can now deploy @MiniMax_AI Speec

We are also launching Science Skills, a specialized bundle that integrates insights from 30+ major life science models and databases with ag…

AgentsDGX agent

We are also launching Science Skills, a specialized bundle that integrates insights from 30+ major life science models and databases with agentic platforms like @Antigravity to allow researchers to pe

working on some hex dashboards -- this is a bummer! if you use langchain/deepagents as your agent harness, you can swap models with fallback…

AgentsDGX agent

working on some hex dashboards -- this is a bummer! if you use langchain/deepagents as your agent harness, you can swap models with fallback middleware when your first choice provider is down https://

19 May 2026

Axial-Relation Guided Fusion State Space Model for Optical-Elevation Sensing Image Segmentation

Local AiDGX agent

arXiv:2605.16768v1 Announce Type: new Abstract: Semantic segmentation of multi-source remote sensing images is a fundamental task for Earth observation applications. Existing methods often struggle wi

CanViT: Toward Active-Vision Foundation Models

Local AiDGX agent

arXiv:2603.22570v2 Announce Type: replace Abstract: Active computer vision promises efficient, biologically plausible perception through sequential, localized glimpses, but lacks scalable general-purp

CoLLM-NAS: Collaborative Large Language Models for Efficient Knowledge-Guided Neural Architecture Search

TutorialsDGX agent

arXiv:2509.26037v2 Announce Type: replace Abstract: The integration of Large Language Models (LLMs) with Neural Architecture Search (NAS) has introduced new possibilities for automating the design of

Confidence Geometry Reveals Trace-Level Correctness in Large Language Model Reasoning

ResearchDGX agent

arXiv:2605.16824v1 Announce Type: cross Abstract: Large language models (LLMs) generate not only reasoning text, but also token-level confidence trajectories that record how uncertainty evolves during

Counterparty Modeling is Not Strategy: The Limits of LLM Negotiators

ResearchDGX agent

arXiv:2605.16575v1 Announce Type: new Abstract: Negotiation requires more than inferring what the other side wants: it requires using that information to make advantageous offers and counteroffers ove

customers are increasingly asking us for certainty on capacity. as models get better, we expect that the world will be capacity-constrained …

IndustryDGX agent

customers are increasingly asking us for certainty on capacity. as models get better, we expect that the world will be capacity-constrained for some time. we are offering discounted tokens for 1-3 yea

Does Your Reasoning Model Implicitly Know When to Stop Thinking?

ResearchDGX agent

arXiv:2602.08354v5 Announce Type: replace Abstract: Recent advancements in large reasoning models (LRMs) have greatly improved their capabilities on complex reasoning tasks through Long Chains of Thou

DriveMoE: Mixture-of-Experts for Vision-Language-Action Model in End-to-End Autonomous Driving

AgentsDGX agent

arXiv:2505.16278v2 Announce Type: replace-cross Abstract: End-to-end autonomous driving (E2E-AD) demands effective processing of multi-view sensory data and robust handling of diverse and complex driv

Embracing Anisotropy: Turning Massive Activations into Interpretable Control Knobs for Large Language Models

ResearchDGX agent

arXiv:2603.00029v2 Announce Type: replace Abstract: Large Language Models (LLMs) exhibit highly anisotropic internal representations, often characterized by massive activations, a phenomenon where a s

Finding Sense in Nonsense with Generated Contexts: Perspectives from Humans and Language Models

ResearchDGX agent

arXiv:2602.11699v3 Announce Type: replace Abstract: Nonsensical and anomalous sentences have been instrumental in the development of computational models of semantic interpretation. A core challenge i

Flow-Direct: Feedback-Efficient and Reusable Guidance for Flow Models via Non-Parametric Guidance Field

ResearchDGX agent

arXiv:2605.16348v1 Announce Type: cross Abstract: Training-free guidance enables pre-trained diffusion and flow models to optimize application-specific objectives using feedback from external black-bo

Forgetting is Competition: Rethinking Unlearning as Representation Interference in Diffusion Models

SafetyDGX agent

arXiv:2603.00975v2 Announce Type: replace-cross Abstract: Deployed text-to-image diffusion models increasingly require post-hoc concept unlearning for copyright claims, artist opt-outs, safety updates

From Demographics to Survey Anchors: Evaluating LLM Agents for Modeling Retirement Attitudes

SafetyDGX agent

arXiv:2605.16303v1 Announce Type: cross Abstract: Large language models (LLM) agents may offer tools to predict human responses to surveys. A common technique for defining these agents uses only demog

It is really good at instruction following, obviously, and having a smart model capable of doing video expands what you can do by a consider…

ApplicationsDGX agent

Ethan Mollick discusses how a model's instruction-following capabilities and video processing abilities expand its potential applications and use cases. He emphasizes that combining strong instruction

Language Acquisition Device in Large Language Models

ResearchDGX agent

arXiv:2605.16758v1 Announce Type: new Abstract: Large Language Models (LLMs) remain substantially less data-efficient than humans. Pre-pretraining (PPT) on synthetic languages has been proposed to clo

Learning What Evaluators Value: A Reliable Approach to Modeling Evaluator Preferences

TutorialsDGX agent

arXiv:2605.16615v1 Announce Type: new Abstract: In many applications, human and LLM evaluators use assessments of relevant criteria to create an overall evaluation for an item or individual. For examp

Leveraging Graph Structure in Seq2Seq Models for Knowledge Graph Link Prediction

ResearchDGX agent

arXiv:2605.18211v1 Announce Type: cross Abstract: We introduce Graph-Augmented Sequence-to-Sequence (GA-S2S), a novel framework that integrates a T5-small encoder-decoder with a Relational Graph Atten

Lotus-2: Advancing Geometric Dense Prediction with Powerful Image Generative Model

ResearchDGX agent

arXiv:2512.01030v3 Announce Type: replace Abstract: Recovering pixel-wise geometric properties from a single image is fundamentally ill-posed due to appearance ambiguity and non-injective mappings bet

Mitigating 3D Prostate Biparametric MRI Data Scarcity through Domain Adaptation using Locally-Trained Latent Diffusion Models for Prostate Cancer Detection

ResearchDGX agent

arXiv:2507.06384v2 Announce Type: replace-cross Abstract: Objective: Latent diffusion models (LDMs) could mitigate data scarcity challenges affecting machine learning development for medical image int

MoleCode unlocks structural intelligence in large language models

ResearchDGX agent

arXiv:2605.16480v1 Announce Type: cross Abstract: Molecules are graphs, but large language models~(LLMs) are usually asked to reason about them through linear strings. The most popular molecular repre

Need advice: Best ComfyUl workflow for texturing a 3D model from 4 orthographic views using reference images?

Local AiDGX agent

A Reddit user seeks guidance on configuring ComfyUI workflows to texture 3D models using orthographic reference images from multiple angles. The discussion likely covers best practices for using Stabl

On the Adversarial Robustness of Large Vision-Language Models under Visual Token Compression

SafetyDGX agent

arXiv:2601.21531v2 Announce Type: replace-cross Abstract: Visual token compression is widely used to accelerate large vision-language models (LVLMs) by pruning or merging visual tokens, yet its advers

One Model to Translate Them All: Universal Any-to-Any Translation for Heterogeneous Collaborative Perception

AgentsDGX agent

arXiv:2605.17907v1 Announce Type: cross Abstract: By sharing intermediate features, collaborative perception extends each agent's sensing beyond standalone limits, but real-world feature modality hete

Peak-Detector: Explainable Peak Detection via Instruction-Tuned Large Language Models in Physiological Sign

SafetyDGX agent

arXiv:2605.16452v1 Announce Type: cross Abstract: Accurate peak detection across diverse cardiac physiological signals, including the Electrocardiogram (ECG), Photoplethysmogram (PPG), Ballistocardiog

Perovskite-R1: a domain-specialized large language model for intelligent discovery of precursor additives and experimental design

ApplicationsDGX agent

arXiv:2507.16307v2 Announce Type: replace-cross Abstract: Perovskite solar cells (PSCs) have rapidly emerged as a leading contender in next-generation photovoltaic technologies, owing to their excepti

PersonaArena: Dynamic Simulation for Evaluating and Enhancing Persona-Level Role-Playing in Large Language Models

AgentsDGX agent

arXiv:2605.17044v1 Announce Type: new Abstract: Large language models (LLMs) increasingly serve as interactive social agents, yet their ability to maintain coherent and authentic persona-level role-pl

Pocket Foundation Models: Distilling TFMs into CPU-Ready Gradient-Boosted Trees

HardwareDGX agent

arXiv:2605.18654v1 Announce Type: cross Abstract: A fraud scorer needs to answer in under 2 ms. The best tabular foundation models (TFMs) take 151-1,275 ms on GPU. We close this gap by distilling the

SG-CADVLM: A Context-Aware Decoding Powered Vision Language Model for Safety-Critical Scenario Generation

SafetyDGX agent

arXiv:2601.18442v3 Announce Type: replace Abstract: Autonomous Vehicle (AV) requires rigorous testing in safety-critical scenarios for safety validation, yet its validation is hindered by the high cos

Sketch Then Paint: Hierarchical Reinforcement Learning for Diffusion Multi-Modal Large Language Models

Local AiDGX agent

arXiv:2605.16842v1 Announce Type: new Abstract: Diffusion Multi-Modal Large Language Models (dMLLMs) are powerful for image generation, but optimizing them through reinforcement learning (RL) remains

Sources: Zyphra, which trains and runs inference for its open-weight models on AMD hardware, is raising a 500M Series B at a valuation of at least 5B (Anna Tong/Forbes)

HardwareDGX agent

Anna Tong / Forbes: Sources: Zyphra, which trains and runs inference for its open-weight models on AMD hardware, is raising a 500M Series B at a valuation of at least 5B — The Series B round, which ch

SparseSAM: Structured Sparsification of Activations in Segment Anything Models

ResearchDGX agent

arXiv:2605.17633v1 Announce Type: cross Abstract: The Segment Anything Model (SAM) achieves strong open-vocabulary segmentation, but its ViT-based image encoders dominate inference latency and memory.

StructLens: A Structural Lens for Language Models via Maximum Spanning Trees

ResearchDGX agent

arXiv:2603.03328v2 Announce Type: replace-cross Abstract: Language exhibits inherent structures, a property that explains both language acquisition and language change. Given this characteristic, we e

The Lattice Representation Hypothesis of Large Language Models

ResearchDGX agent

arXiv:2603.01227v2 Announce Type: replace Abstract: We propose the Lattice Representation Hypothesis of large language models: a symbolic backbone that grounds conceptual hierarchies and logical opera

Towards Generalized Image Manipulation Localization via Score-based Model

Local AiDGX agent

arXiv:2605.16879v1 Announce Type: new Abstract: With the rapid evolution of synthetic media, Image Manipulation Localization (IML) has emerged as a critical component in multimedia forensics for ensur

Trajectory-Aware Adaptive Inference in Object Detection Models

AgentsDGX agent

arXiv:2605.16397v1 Announce Type: cross Abstract: The increasing integration of sensors in autonomous maritime navigation has led to large-scale multimodal datasets, raising challenges in achieving ef

UB-SMoE: Universally Balanced Sparse Mixture-of-Experts for Resource-adaptive Federated Fine-tuning of Foundation Models

ResearchDGX agent

arXiv:2605.16690v1 Announce Type: new Abstract: Heterogeneous LoRA-rank methods address system heterogeneity in federated fine-tuning of foundation models by assigning client-specific ranks based on c

VLM-AutoDrive: Post-Training Vision-Language Models for Safety-Critical Autonomous Driving Events

SafetyDGX agent

arXiv:2603.18178v2 Announce Type: replace-cross Abstract: The rapid growth of ego-centric dashcam footage presents a major challenge for detecting safety-critical events such as collisions and near-co

Wasserstein bounds for denoising diffusion probabilistic models via the Follmer process

ResearchDGX agent

arXiv:2605.18069v1 Announce Type: cross Abstract: This paper studies sampling error bounds for denoising diffusion probabilistic models (DDPMs) in the 2-Wasserstein distance. Our contributions are thr

We just added Grok's new imagine model in Paper so you can explore images even faster. Here's what we found: - Super fast generations for th…

IndustryDGX agent

We just added Grok's new imagine model in Paper so you can explore images even faster. Here's what we found: - Super fast generations for the quality - Saves details when editing - Perfect for 30 rapi

What Matters for Grocery Product Retrieval with Open Source Vision Language Models

ResearchDGX agent

arXiv:2605.18029v1 Announce Type: new Abstract: Multimodal product retrieval (MPR) underpins checkout-free retail and automated inventory systems, yet it demands fine-grained SKU discrimination that s

When Dynamics Shift, Robust Task Inference Wins: Offline Imitation Learning with Behavior Foundation Models Revisited

SafetyDGX agent

arXiv:2605.17017v1 Announce Type: cross Abstract: Behavior Foundation Models (BFMs) enable scalable imitation learning (IL) by pretraining task-agnostic representations that can be rapidly adapted to

18 May 2026

3 new models from @xai's Grok creative stack are live on OpenRouter: • Grok Imagine Image Quality: photoreal image generation and editing • …

IndustryDGX agent

3 new models from @xai's Grok creative stack are live on OpenRouter: • Grok Imagine Image Quality: photoreal image generation and editing • Grok Imagine Video: short clips from text, image, or referen

A mental model for working with coding agents is that they're blind squirrels running into a maze and bumping into walls. You must place the…

ResearchDGX agent

A mental model for working with coding agents is that they're blind squirrels running into a maze and bumping into walls. You must place the walls (verifiable constraints) strategically so that they e

← Previous
1…216217218219220…1017
Next →