AI Wiki
TimelineEvolutionGraphStatusAsk wiki
Live from Git
AI Wiki
TimelineEvolutionGraphStatusAsk wiki
Live from Git
Filter entries
Categories
  • All entries85,115
  • Agents7,313
  • Applications5,228
  • Concepts5
  • Hardware1,762
  • Industry6,105
  • Local Ai4,756
  • Model Releases22,759
  • Research19,333
  • Safety12,889
  • Syntheses17
  • Tools1,669
  • Tutorials3,279

Source
HumanDGX agent

Content type
AllBlogX PostPaperYouTubeRedditGitHub
Clear filters
Categories
  • All entries85,115
  • Agents7,313
  • Applications5,228
  • Concepts5
  • Hardware1,762
  • Industry6,105
  • Local Ai4,756
  • Model Releases22,759
  • Research19,333
  • Safety12,889
  • Syntheses17
  • Tools1,669
  • Tutorials3,279

Source
HumanDGX agent

85,115Total entries
1Added by human
85,114Found by agent
12Categories

Knowledge catalogue

Search: “models”

GridTimelineEvolution
60,986 results
1 Jul 2026

Mean-Field Model for Two-Layer Neural Networks Trained with Consensus-Based Optimization

ResearchDGX agent

arXiv:2511.21466v3 Announce Type: replace Abstract: We study Consensus-Based Optimization (CBO) for two-layer neural network training. We compare the performance of CBO against Adam on two test cases

Model subsidies are ending. What do you do now?

AgentsDGX agent

Flat-rate AI plans are subsidizing agentic workloads. Learn why LLM inference costs are moving to metered pricing and how evals reveal cost per successful task. The post Model subsidies are ending. Wh

Nano Banana 2 Lite is now in Comfy. The fastest, cheapest model in the Nano Banana family. → 4-second image gen → $0.034 per 1K images If yo…

Local AiDGX agent

Nano Banana 2 Lite is now in Comfy. The fastest, cheapest model in the Nano Banana family. → 4-second image gen → $0.034 per 1K images If your pipeline runs on iteration volume, this changes the game.

Content type
AllBlogX PostPaperYouTubeRedditGitHub
Clear filters

On Optimizing Multimodal Jailbreaks for Spoken Language Models

SafetyDGX agent

arXiv:2603.19127v2 Announce Type: replace Abstract: As Spoken Language Models (SLMs) integrate speech and text modalities, they inherit the safety vulnerabilities of their LLM backbone while introduci

Region-Aware Multimodal Large Language Model via SlowFast Tokenization and Pseudo-Mask Guidance for 3D CT Report Generation

TutorialsDGX agent

arXiv:2506.23102v2 Announce Type: replace-cross Abstract: Current CT report generation frameworks predominantly rely on global feature representations, often failing to capture region-specific details

Robust Text Watermarking for Large Language Models via Dual Semantic Embeddings

SafetyDGX agent

arXiv:2606.31602v1 Announce Type: new Abstract: This work presents Dual-Embedding Watermarking (DEW), a semantic watermarking scheme for large language models (LLMs) that leverages contextual and toke

SAGE: A Search-AuGmented Evaluation of Large Language Models on Free-Form QA

AgentsDGX agent

arXiv:2504.07385v3 Announce Type: replace-cross Abstract: As Large Language Models (LLMs) become increasingly used for question-answering (QA), relying on static, pre-annotated references for evaluati

Twelve Labs, which is building AI models to make video searchable and understandable, raised a $100M Series B co-led by NEA and Naver, and signs an AWS deal (Saritha Rai/Bloomberg)

IndustryDGX agent

Saritha Rai / Bloomberg: Twelve Labs, which is building AI models to make video searchable and understandable, raised a 100M Series B co-led by NEA and Naver, and signs an AWS deal — Twelve Labs Inc.

Venice AI, which offers access to 200+ AI models while allowing users to retain their privacy, raised a 65M Series A led by Dragonfly at a 1B valuation (Ram Iyer/TechCrunch)

SafetyDGX agent

Ram Iyer / TechCrunch: Venice AI, which offers access to 200+ AI models while allowing users to retain their privacy, raised a 65M Series A led by Dragonfly at a 1B valuation — Concerns over the impac

Wavelet-Optimized Pseudo-3D Accelerated Diffusion Model for Truncated Computed Laminography

ApplicationsDGX agent

arXiv:2606.31318v1 Announce Type: new Abstract: Computed Laminography (CL) is a key technology for the nondestructive testing of large plate-shaped objects. However, field-of-view (FOV) limitations in

Wind and State Estimation on SE(3): Comparative Evaluation of EKF and UKF with Continuous and Discrete Quadrotor Models

ResearchDGX agent

arXiv:2606.30804v1 Announce Type: new Abstract: Use of quadrotor UAVs for wind velocity estimation is gaining popularity in recent studies, leveraging their maneuverability, compact size and low cost.

You can start building with GLM 5.2 in minutes: 1️⃣ Download dcode 2️⃣ Select your model (GLM 5.2) 3️⃣ Enter your API key …and you're ready …

AgentsDGX agent

You can start building with GLM 5.2 in minutes: 1️⃣ Download dcode 2️⃣ Select your model (GLM 5.2) 3️⃣ Enter your API key …and you're ready to go with frontier performance on open weights. Great demo

ZEBRA: Zero-Shot Entropy-Regularized Prompt Learning for Base-to-Novel Generalization in Audio-Language Models

ResearchDGX agent

arXiv:2606.31587v1 Announce Type: cross Abstract: Audio-Language Models (ALMs) achieve strong zero-shot performance by aligning audio with textual class descriptions. Although prompt learning improves

30 Jun 2026

Adam's Law: Textual Frequency Law on Large Language Models

AgentsDGX agent

arXiv:2604.02176v3 Announce Type: replace Abstract: While textual frequency has been validated as relevant to human cognition in reading speed, its relatedness to Large Language Models (LLMs) is seldo

After 18 months of writing, coding, and experimenting, Build a Reasoning Model (From Scratch) is finally out! My first copies just arrived! …

ResearchDGX agent

After 18 months of writing, coding, and experimenting, Build a Reasoning Model (From Scratch) is finally out! My first copies just arrived! 📚 440 full-color pages. Inference scaling, reinforcement lea

Agentic Tool Use in Large Language Models

SafetyDGX agent

arXiv:2604.00835v2 Announce Type: replace Abstract: Large language models are increasingly being deployed as autonomous agents yet their real world effectiveness depends on reliable tools for informat

Analyzing Uncertainty in the Spatial Representation of the Kinematic Bicycle Model

AgentsDGX agent

arXiv:2606.29566v1 Announce Type: new Abstract: Locating a vehicle and determining its orientation in an uncertain environment is a critical challenge in autonomous vehicle navigation and path plannin

'At this very moment China is giving its AI technology away. It's releasing open-weight AI models that are cheap, capable, and they're fast …

IndustryDGX agent

'At this very moment China is giving its AI technology away. It's releasing open-weight AI models that are cheap, capable, and they're fast becoming the world's default.' We can overcome this. @neil_c

Bandwidth Selection in Kernel Density Estimation for Model Calibration

SafetyDGX agent

arXiv:2606.29925v1 Announce Type: new Abstract: As deep learning models are increasingly deployed in high-stakes applications, providing well-calibrated uncertainty estimates has become as critical as

Characterizing Large Language Model Agentic Workflows: A Study on N8n Ecosystem

SafetyDGX agent

arXiv:2606.29116v1 Announce Type: new Abstract: Large Language Models (LLMs) are rapidly being adopted in low-code and no-code automation platforms, where non-expert users design workflows that combin

CLMASP: Coupling Large Language Models with Answer Set Programming for Robotic Task Planning

ResearchDGX agent

arXiv:2406.03367v2 Announce Type: replace Abstract: Large Language Models (LLMs) possess extensive foundational knowledge and moderate reasoning abilities, making them suitable for general task planni

COMPASS: Grounding Composition-Intent Guidance in Unified Multimodal Models

ResearchDGX agent

arXiv:2606.28696v1 Announce Type: new Abstract: Composition is a high-level visual intent that governs where subjects are placed and how a scene is organized, yet current unified multimodal models rem

Connectivity Estimation using Stochastic Graph Heat Modelling

ApplicationsDGX agent

arXiv:2606.29098v1 Announce Type: cross Abstract: A growing number of techniques leverage the spatial structures that underlie many real-world datasets. Despite these advances, the complementary task

Conversational Domain Adaptation of IndicTrans2 across 21 Indic Languages via Experience Replay and Model Soups

ResearchDGX agent

arXiv:2606.29024v1 Announce Type: new Abstract: IndicTrans2 is the strongest open English to Indic translation system, but like most systems it is trained on general text and tends to sound stiff on c

ENC-ODE: Event-level Neurodegenerative Modeling in Continuous Time with Neural ODEs

ResearchDGX agent

arXiv:2606.30398v1 Announce Type: new Abstract: Accurately predicting the temporal evolution of clinical biomarkers is crucial for the early diagnosis and management of neurodegenerative diseases such

Enhanced Diffusion Sampling: Efficient Rare Event Sampling and Free Energy Calculation with Diffusion Models

HardwareDGX agent

arXiv:2602.16634v2 Announce Type: replace-cross Abstract: The rare-event sampling problem has long been the central limiting factor in molecular dynamics (MD), especially in biomolecular simulation. R

Evolutionary Hyperparameter Optimization to Find Lightweight CNN Models for Autonomous Steering

AgentsDGX agent

arXiv:2606.29684v1 Announce Type: cross Abstract: This research investigates the optimization of Convolutional and Dense Neural Networks (CNNs and DNNs) for autonomous steering using the (N+M) Evoluti

FADE: Mitigating Hallucinations by Reducing Language-Prior Dominance in Large Vision-Language Models

ResearchDGX agent

arXiv:2606.29431v1 Announce Type: new Abstract: Despite the impressive capabilities of Large Vision-Language Models (LVLMs), they remain susceptible to hallucination, generating content inconsistent w

Fast Wireless Foundation Models with Early-Exits

ResearchDGX agent

arXiv:2606.29640v1 Announce Type: cross Abstract: While wireless foundation models (FMs) are demonstrating strong potential to enable AI-Native 6G networks, their high computational cost remains a cri

GeNeRT: A Physics-Informed Approach to Intelligent Wireless Channel Modeling via Generalizable Neural Ray Tracing

TutorialsDGX agent

arXiv:2506.18295v2 Announce Type: replace-cross Abstract: Neural ray tracing (RT) has emerged as a promising paradigm for channel modeling by integrating physical propagation principles with neural ne

Invariant Reasoning Directions in Latent Trajectories of Language Models

SafetyDGX agent

arXiv:2606.29164v1 Announce Type: cross Abstract: Latent reasoning models perform multi-step inference directly in hidden-state space, yet the structure of these latent reasoning trajectories remains

LeVo 2: Stable and Melodious Song Generation via Hierarchical Representation Modeling and Progressive Post-Training

SafetyDGX agent

arXiv:2606.30642v1 Announce Type: cross Abstract: Full-length song generation must preserve coherence and musicality, render detailed vocal and accompaniment acoustics, and follow lyrics and prompts.

LoRA-Tuned Large Language Models for Dementia Detection via Multi-View Speech-Derived Features

TutorialsDGX agent

arXiv:2606.28445v1 Announce Type: cross Abstract: Early detection of dementia enables timely intervention, and reflecting cognitive impairment, spontaneous speech offers a non-invasive screening modal

Modelling Human Values for Value-Aware Multi-Agent Systems

SafetyDGX agent

arXiv:2402.06359v2 Announce Type: replace Abstract: One of today's most pressing societal challenges is building AI systems whose behaviour, or the behaviour it enables within communities of interacti

Notes on generative modeling: flow matching, diffusion, optimal transport and Schr{odinger bridge

ResearchDGX agent

arXiv:2606.30053v1 Announce Type: cross Abstract: These notes recapitulate the high level mathematical principles behind different techniques for generative modeling. I show the connections between op

On the Faithfulness of Post-Hoc Concept Bottleneck Models

ApplicationsDGX agent

arXiv:2606.30498v1 Announce Type: cross Abstract: Human decision-making interprets the world through high-level concepts, such as recognizing a bird by its belly color. To bridge the gap between opaqu

Ontology-Guided Reverse Thinking Makes Large Language Models Stronger on Knowledge Graph Question Answering

TutorialsDGX agent

arXiv:2502.11491v3 Announce Type: replace-cross Abstract: Large language models (LLMs) have shown remarkable capabilities in natural language processing. However, in knowledge graph question answering

OpenSPM: An Environment-Transferable Robotic Key Spatial Pose Memory and Closed-Loop High-Frequency Flow-Matching Action Generation Model

ResearchDGX agent

arXiv:2606.29936v1 Announce Type: new Abstract: Open-environment tabletop robotic manipulation requires systems to possess semantic understanding, precise geometric pose estimation, and high-frequency

OWMDrive: Causality-Aware End-to-End Autonomous Driving via 4D Occupancy World Model

SafetyDGX agent

arXiv:2606.30421v1 Announce Type: new Abstract: Autonomous driving systems are steadily moving toward end-to-end paradigms to mitigate the limited adaptability of rule-based pipelines in complex traff

Robustness and Structure Preservation in Flow-Based Generative Models via Wasserstein Path-Space Divergences

SafetyDGX agent

arXiv:2410.01244v2 Announce Type: replace-cross Abstract: We introduce a novel Wasserstein-1 (W_1) path-space divergence for stochastic and deterministic dynamics and establish a Wasserstein Uncertain

SATB-VR: Training Few-Step Video Restoration Diffusion Model using SNR-Aware Trajectory Blending

ApplicationsDGX agent

arXiv:2606.28677v1 Announce Type: new Abstract: While diffusion models excel in video restoration, their reliance on extensive iterative steps limits efficiency. Conversely, aggressive single-step dis

Set-Inclusive Uncertainty Modeling for Robust Brain Tumor Segmentation

ResearchDGX agent

arXiv:2606.30374v1 Announce Type: cross Abstract: Multimodal MRI is essential for accurate brain tumor segmentation. However, acquiring all modalities at inference is often challenging in practice, wh

Simplify multi-account access to Amazon Bedrock models with managed entitlements

TutorialsDGX agent

In this post, we show you how to use managed entitlements for Amazon Bedrock to subscribe once from a central account and distribute model access across your organization. This approach removes the ne

Understanding Evaluation Illusion in Diffusion Large Language Models

ResearchDGX agent

arXiv:2606.29228v1 Announce Type: new Abstract: Despite the capability of parallel decoding, diffusion large language models (dLLMs) require many denoising steps to maintain generation quality, motiva

Using Large Language Models as Low-Cost Statistical Estimators for Human-Response Data

SafetyDGX agent

arXiv:2606.30372v1 Announce Type: new Abstract: Quantitative research across the social and behavioral sciences depends on human subject experiments that are expensive, slow, and subject to sampling b

29 Jun 2026

Bifocal Diffusion Language Models: Asymmetric Bidirectional Context for Parallel Generation

ResearchDGX agent

arXiv:2606.27732v1 Announce Type: cross Abstract: Discrete diffusion language models (dLLMs) recover masked tokens in parallel, offering significant speedups over autoregressive (AR) generation. Howev

CacheMPC: Certified Cached Model Predictive Control for Quadruped Locomotion

SafetyDGX agent

arXiv:2606.28300v1 Announce Type: new Abstract: Model Predictive Control (MPC) is the standard predictive layer in hierarchical quadruped controllers, but the per-cycle QP solve limits the update rate

If you are building harnesses or AI tools and would like help adding grok in there hit me up More powerful models are coming and it's better…

IndustryDGX agent

Elon Musk offers assistance to developers building harnesses or AI tools who want to integrate Grok, Musk's AI model, into their applications. He indicates that more advanced versions of Grok are in d

IWP: Token Pruning as Implicit Weight Pruning in Large Vision Language Models

ResearchDGX agent

arXiv:2604.00757v2 Announce Type: replace-cross Abstract: Large Vision Language Models show impressive performance across image and video understanding tasks, yet their computational cost grows rapidl

Panoramic Scene Analysis: A Survey from Distortion-Aware Engineering to Sphere-Native Foundation Modeling

ResearchDGX agent

arXiv:2606.27745v1 Announce Type: new Abstract: Panoramic images capture the complete visual sphere in a single frame, providing spatial context unattainable by conventional cameras. Yet this complete

POTracker: Optimizing Large Language Models for Standard-Compliant Power Outage Report Generation

ResearchDGX agent

arXiv:2606.23533v2 Announce Type: replace Abstract: Recent large language models (LLMs) are good at general text generation, but it is still hard to use them for domain-specific data generation becaus

Recall Before Rerank: Benchmarking Deep Learning Models for Large-Scale Code-to-Code Retrieval

ResearchDGX agent

arXiv:2606.27401v1 Announce Type: cross Abstract: Semantic code search and clone detection are essential for software development, maintenance, and reuse. This paper evaluates the effectiveness, effic

SITUATION EXPLAINED: Why is restricting open source AI models structurally impossible? @ClementDelangue, co-founder and CEO of @huggingface:…

SafetyDGX agent

SITUATION EXPLAINED: Why is restricting open source AI models structurally impossible? @ClementDelangue, co-founder and CEO of @huggingface: 'Open weights are fundamentally different than an API. The

Textual Belief States for World Models: Identifiable Representation Learning Under Strict Mediation

TutorialsDGX agent

arXiv:2606.27681v1 Announce Type: cross Abstract: World models in partially observed environments rely on latent representations that summarize interaction history, but in many modern LLM-based archit

Today, we are releasing Rampart: a 14.7MB machine learning model designed to protect citizens’ privacy by redacting personal information dir…

IndustryDGX agent

Today, we are releasing Rampart: a 14.7MB machine learning model designed to protect citizens’ privacy by redacting personal information directly in your browser before it gets sent to any server A bi

Try SpaceXAI Voice models in the Vercel AI Gateway

IndustryDGX agent

Try SpaceXAI Voice models in the Vercel AI Gateway Grok's realtime voice is now on AI Gateway. Build with AI SDK 7: • 𝚡𝚊𝚒/𝚐𝚛𝚘𝚔-𝚟𝚘𝚒𝚌𝚎-𝚝𝚑𝚒𝚗𝚔-𝚏𝚊𝚜𝚝-𝟷.𝟶 (𝚞𝚜𝚎𝚁𝚎𝚊𝚕𝚝𝚒𝚖𝚎) • 𝚡𝚊𝚒/𝚐𝚛𝚘𝚔-𝚝𝚝𝚜 (𝚐𝚎𝚗𝚎𝚛𝚊𝚝𝚎𝚂𝚙𝚎𝚎𝚌𝚑) • 𝚡𝚊𝚒/

28 Jun 2026

There's a big difference between a single model call and serving an agent at scale. @ZainHasan6 breaks down what actually changes. Catch our…

AgentsDGX agent

There's a big difference between a single model call and serving an agent at scale. @ZainHasan6 breaks down what actually changes. Catch our team this Monday at 9 a.m. PST for their open-source infere

27 Jun 2026

Every enterprise will have its own model-harness-sandbox-eval flywheel with token value per watt optimization. This is the future. Simple re…

ApplicationsDGX agent

Every enterprise will have its own model-harness-sandbox-eval flywheel with token value per watt optimization. This is the future. Simple reason: tacit knowledge about the domain and customers and the

not dunking on Dario here bc Mythos is to GPT2 what a lion is to a bumblebee, but it’s insane that GPT-2, and successive models, have had ~0…

SafetyDGX agent

not dunking on Dario here bc Mythos is to GPT2 what a lion is to a bumblebee, but it’s insane that GPT-2, and successive models, have had ~0% impact on causing bad universal outcomes & so SHOULD HAVE

Tokyo-based Sakana AI's Fugu and China-based 360's cybersecurity model Tulongfeng claim to rival Anthropic's banned Mythos and Fable 5 amid the US export ban (Kate Park/TechCrunch)

IndustryDGX agent

Kate Park / TechCrunch: Tokyo-based Sakana AI's Fugu and China-based 360's cybersecurity model Tulongfeng claim to rival Anthropic's banned Mythos and Fable 5 amid the US export ban — On Wednesday, Ch

← Previous
1…207208209210211…1017
Next →