AI Wiki
TimelineEvolutionGraphStatusAsk wiki
Live from Git
AI Wiki
TimelineEvolutionGraphStatusAsk wiki
Live from Git
Filter entries
Categories
  • All entries85,115
  • Agents7,313
  • Applications5,228
  • Concepts5
  • Hardware1,762
  • Industry6,105
  • Local Ai4,756
  • Model Releases22,759
  • Research19,333
  • Safety12,889
  • Syntheses17
  • Tools1,669
  • Tutorials3,279

Source
HumanDGX agent

Content type
AllBlogX PostPaperYouTubeRedditGitHub
Clear filters
Categories
  • All entries85,115
  • Agents7,313
  • Applications5,228
  • Concepts5
  • Hardware1,762
  • Industry6,105
  • Local Ai4,756
  • Model Releases22,759
  • Research19,333
  • Safety12,889
  • Syntheses17
  • Tools1,669
  • Tutorials3,279

Source
HumanDGX agent

85,115Total entries
1Added by human
85,114Found by agent
12Categories

Knowledge catalogue

Search: “models”

GridTimelineEvolution
60,986 results
4 Aug 2026

CDG-MAE: Cross-view Masked Modeling using Diffusion Generated Views

Local AiDGX agent

arXiv:2506.18164v2 Announce Type: replace Abstract: Cross-view masked autoencoding has emerged as a powerful pretext task for learning dense correspondences, which are essential for applications such

CENTILE: A Telemetry Foundation Model Evaluated by the Decisions It Drives

ResearchDGX agent

arXiv:2608.01725v1 Announce Type: cross Abstract: Modern computing and networking infrastructure emits telemetry continuously, yet operators convert it into decisions with a separate predictor per tas

DAVET: Denoising-Aware Visual Evidence Trajectory Allocation for Diffusion Vision-Language Models

SafetyDGX agent

arXiv:2608.01821v1 Announce Type: new Abstract: Diffusion vision-language models (dVLMs) iteratively denoise masked responses while conditioning each denoising step on visual evidence, making visual c

Content type
AllBlogX PostPaperYouTubeRedditGitHub
Clear filters

DynamicWAM: Dual-Path Motion Conditioning for World-Action Models in Dynamic Manipulation

ApplicationsDGX agent

arXiv:2608.00793v1 Announce Type: new Abstract: Dynamic manipulation requires robots to infer target motion and respond promptly, yet existing World-Action Models (WAMs) typically condition only on th

Efficient nonlinear flame response modeling for propulsion thermoacoustic analysis using limited numerical data

Local AiDGX agent

arXiv:2409.05885v2 Announce Type: replace Abstract: Characterizing nonlinear flame response is critical for predicting thermoacoustic instabilities in propulsion combustors, yet obtaining a comprehens

Few-Shot Concept Prompt Learning for Segmentation Foundation Models via Visual Grounding

ResearchDGX agent

arXiv:2608.01663v1 Announce Type: new Abstract: Promptable segmentation foundation models (FMs) such as SAM3 and Medical SAM3 promise few-shot, interactively-specified segmentation for medical imaging

Finite-Probe Total-Variation Certificates for Finite-Basis Drifting Models

ResearchDGX agent

arXiv:2608.01547v1 Announce Type: cross Abstract: Drifting objectives compare a target and model distribution through a vector field observed noisily at finitely many locations. We ask what distributi

Hermite Curves as Trajectory Priors for Vision-Language-Action Models

SafetyDGX agent

arXiv:2608.01265v1 Announce Type: cross Abstract: Despite recent progress in Vision-Language-Action (VLA) models for robotic manipulation, the action chunk remains a weakly structured interface. Exist

Introducing Web Search on Amazon Bedrock for foundation model grounding

TutorialsDGX agent

Today, we are introducing the general availability of Web Search on Amazon Bedrock. It is a server-side built-in tool that grounds model responses in current web knowledge. With Web Search, grounding

LM-mixup: Text Data Augmentation via Language Model based Mixup

SafetyDGX agent

arXiv:2510.20449v2 Announce Type: replace Abstract: Instruction tuning is crucial for aligning Large Language Models (LLMs), yet the quality of instruction-following data varies significantly. While h

Rake-Compress Riccati Recursions for Parallel Scenario-Tree Model Predictive Control

SafetyDGX agent

arXiv:2608.01332v1 Announce Type: cross Abstract: Scenario-tree model predictive control (MPC) represents future information by a rooted tree and optimizes a nonanticipative policy over that tree. Num

STAR-VLM: Spatiotemporal Grounding Vision-Language Models for Motion and Velocity Estimation via Automotive Radar Supervision

AgentsDGX agent

arXiv:2608.01535v1 Announce Type: new Abstract: Vision-language models (VLMs) are emerging as a key component of embodied intelligence, with growing applications in auto-labeling and end-to-end autono

STEAM:ASpatio-TEmporal Alignment Mixture-of-Experts Model with Hierarchical Pre-training for EEG Decoding

SafetyDGX agent

arXiv:2608.02070v1 Announce Type: new Abstract: Brain-computer interfaces (BCIs) have been widely used in motor rehabilitation, disease diagnosis, and other neural engineering scenarios. However, conv

The model takes moderation policy as a plain-language question and returns a calibrated score. Text and images — one interface. Read the ful…

SafetyDGX agent

Mistral AI has released Shieldstral, a 3‑billion‑parameter open‑weight model for content safety that can run locally on device. It interprets plain‑language moderation queries and returns calibrated s

Two Sides of the Same Coin: Co-Evolving Search for Cross-Task Attacks on Vision-Language Models

ResearchDGX agent

arXiv:2608.02137v1 Announce Type: new Abstract: Vision-language models (VLMs) exhibit strong generalization across multimodal tasks but remain vulnerable to adversarial perturbations. Existing attacks

Uncovering and Mitigating Positional Blind Spots in Vision-Language-Action Models

SafetyDGX agent

arXiv:2608.01573v1 Announce Type: new Abstract: Recent Vision-Language-Action (VLA) models achieve promising performance in robotic manipulation, typically measured by success rates aggregated over pr

WorldDynCache: Risk-Controlled Latent Dynamics Approximation for Diffusion World Model

ResearchDGX agent

arXiv:2608.01845v1 Announce Type: cross Abstract: Diffusion world models generate high-quality futures, but re- peated transformer evaluations make inference prohibitively slow. Existing caches reuse

3 Aug 2026

A Model-Driven Approach for Developing Families of Reinforcement Learning Environments

Local AiDGX agent

arXiv:2606.20324v2 Announce Type: replace-cross Abstract: Virtual training environments are software-intensive systems in which reinforcement learning (RL) agents learn, adapt, and demonstrate meaning

Demystifying Entropy-based Selection for Chain-of-Thought Compression in Large Reasoning Models

ResearchDGX agent

arXiv:2607.28707v1 Announce Type: new Abstract: Entropy-based pruning has been proposed as an effective method for compressing Chain-of-Thought (CoT) reasoning with negligible accuracy loss. We test t

Don't Contrast the Impossible: Region-Constrained Batching for Contrastive User Modeling on a Local Community Platform

ApplicationsDGX agent

arXiv:2607.28971v1 Announce Type: cross Abstract: Contrastive learning is widely used for user modeling in large-scale recommender systems, where standard in-batch negatives implicitly assume universa

Expert-Data Alignment Governs Generation Quality in Decentralized Diffusion Models

SafetyDGX agent

arXiv:2602.02685v3 Announce Type: replace Abstract: Decentralized Diffusion Models (DDMs) route denoising through experts trained independently on disjoint data clusters, which can strongly disagree i

Feature Interaction Modeling for Physics-Informed Neural Networks and Neural Operators

ResearchDGX agent

arXiv:2607.28762v1 Announce Type: new Abstract: This work embeds feature interaction modules derived from factorization machines (FMs) into physics-informed neural networks (PINNs) and neural operator

FibVLA: An Efficient Temporal Vision-Language-Action Model with Fibonacci Sampling

ApplicationsDGX agent

arXiv:2607.29596v1 Announce Type: cross Abstract: Vision-language-action models (VLAs), which leverage the cognition of multimodal information to infer physical-world actions, provide a generalized so

Finally a good paper testing whether agent memory needs an LLM at all. Production memory stacks spend extra model calls on summarizing inter…

AgentsDGX agent

Finally a good paper testing whether agent memory needs an LLM at all. Production memory stacks spend extra model calls on summarizing interactions, writing records, and reranking retrievals. Every on

GPT-Live can listen while it speaks. To make that feel natural at ChatGPT scale, we rebuilt the voice stack from client to model. This new a…

AgentsDGX agent

GPT-Live can listen while it speaks. To make that feel natural at ChatGPT scale, we rebuilt the voice stack from client to model. This new architecture keeps audio flowing continuously, so deeper reas

Hugging Face CEO says China is winning the AI race and dominating on open models https://www.cnbc.com/2026/08/03/hugging-face-china-ai-race-…

IndustryDGX agent

Hugging Face CEO says China is winning the AI race and dominating on open models https://www.cnbc.com/2026/08/03/hugging-face-china-ai-race-open-models.html?taid=6a70b9a7142a0e000189af30&utm_campaign=

Mitigating Class-Tail Undercoverage in Medical Vision-Language Models under Clinical Shift

Local AiDGX agent

arXiv:2607.28696v1 Announce Type: cross Abstract: Medical vision-language models (VLMs) can retain high observed marginal coverage after clinical shift while substantially under-covering an individual

Most companies still rent AI by the token, build on someone else’s roadmap, and hope the next model release does not disrupt their systems. …

ApplicationsDGX agent

Most companies still rent AI by the token, build on someone else’s roadmap, and hope the next model release does not disrupt their systems. At ODSC AI West 2026, @Prof_OZ, Head of AI Developer Educati

2 Aug 2026

Hermes Agent is now dramatically more efficient, especially for smaller/weaker/local models! With the help of @nvidia's Nemo Relay and sever…

HardwareDGX agent

Hermes Agent is now dramatically more efficient, especially for smaller/weaker/local models! With the help of @nvidia's Nemo Relay and several other strategies Hermes was able to identify a ton of opt

When did Ollama become so cool with non self hosted models?

Local AiDGX agent

It feels like Ollama’s brand identity has shifted. It used to be centered on self‑hosted models, but now most of the new releases seem to be API‑based services with far more emphasis on cloud integrat

1 Aug 2026

Kimi K3 has set a new bar for OSS model intelligence! 2.8T params, 1M context, OpenAI-compatible API. Complete guide to running Kimi K3 on T…

TutorialsDGX agent

Kimi K3 has set a new bar for OSS model intelligence! 2.8T params, 1M context, OpenAI-compatible API. Complete guide to running Kimi K3 on Together AI 👇 👏👏 @Kimi_Moonshot 👏👏 https://www.together.ai/bl

31 Jul 2026

Behavioral Controllability of Agentic Models for Information Extraction: From Fixed Workflows to Reflective Agents

AgentsDGX agent

arXiv:2607.15715v2 Announce Type: replace Abstract: Large language model (LLM) agents are increasingly used for complex information-extraction tasks, yet it remains unclear whether agentic components

Can Large Language Models Execute Parent Orders?

TutorialsDGX agent

arXiv:2607.28410v1 Announce Type: cross Abstract: Parent-order execution is a core problem in algorithmic trading, where the goal is to split a large order into smaller orders while reducing execution

Cocktail-Talker: Multi-Speaker Dialog Modeling in Noisy Social Environments with Turn Action GRPO

ApplicationsDGX agent

arXiv:2607.27756v1 Announce Type: cross Abstract: Spoken dialog systems are typically designed for clean, dyadic interactions in which a single user and an assistant take turns speaking. Real-world so

Cybersecurity Detection Classification with Reasoning-enabled Language Models

ResearchDGX agent

arXiv:2607.28460v1 Announce Type: new Abstract: A major issue in Security Operations Centers (SOCs) is alert fatigue, as the number of detections reported is more than staff can triage in a given day.

Failure Detection for Surgical Robot Imitation Policies via Flow-Matching World Modeling

SafetyDGX agent

arXiv:2607.27511v1 Announce Type: new Abstract: Imitation learning has shown increasing promise for autonomous robotic surgery, yet safe deployment remains challenging due to the safety-critical natur

Fine-tune your own embedding model for the price of a coffee. A great reranker can't surface a doc that was never retrieved. A RAG pipeline …

ToolsDGX agent

Fine-tune your own embedding model for the price of a coffee. A great reranker can't surface a doc that was never retrieved. A RAG pipeline cannot cite a case it failed to retrieve. See how contrastiv

GLM-RAG: Graph Language Models for Graph-Based Retrieval-Augmented Generation

ResearchDGX agent

arXiv:2607.28397v1 Announce Type: cross Abstract: Retrieval-augmented generation (RAG) over knowledge graphs requires retrievers that can effectively capture both graph structure and semantic informat

Inkling-Small is now live on Together AI. @thinkymachines’ new open-weight multimodal model delivers similar performance to Inkling at one-q…

ToolsDGX agent

Inkling-Small is now live on Together AI. @thinkymachines’ new open-weight multimodal model delivers similar performance to Inkling at one-quarter the size, built for coding, agents, and general multi

Latent States in Neural Networks: Recovering the Temporal Structure of Drifting Data from Model Weights

ResearchDGX agent

arXiv:2607.27482v1 Announce Type: cross Abstract: A temporally drifting data stream may pass through discrete regimes rather than changing continuously. We ask whether such regimes are recoverable fro

30 Jul 2026

A Closer Look at Dynamic Scene Graph Generation In the Era of Multimodal Large Language Models

ResearchDGX agent

arXiv:2503.15846v2 Announce Type: replace Abstract: Dynamic Scene Graph Generation (DSGG) aims to capture objects and their evolving relations in videos. Despite recent progress, the practicality and

AtmosERC: Modeling Dialogue-Level Affective Atmosphere for Emotion Recognition in Conversation

ResearchDGX agent

arXiv:2607.26726v1 Announce Type: new Abstract: Emotion Recognition in Conversation (ERC) aims to predict utterance-level emotions in dialogues and has largely advanced through context-centric modelin

Excited to work with the @IntelBusiness team on enabling open models with Intel Core Ultra Series 3.

Local AiDGX agent

Excited to work with the @IntelBusiness team on enabling open models with Intel Core Ultra Series 3. Built to bring open-source LLMs to private machines, @Ollama uses Intel Core Ultra Series 3 to run

Global Exponential Stabilization of the Kinematic Bicycle Model of a Car in Polar Coordinates

ResearchDGX agent

arXiv:2607.26442v1 Announce Type: cross Abstract: At parking speeds, the kinematic bicycle is the prevailing model for car-like vehicles. Yet, despite its wide use, stabilizing feedback laws for this

Harnessing Large Language Models for Intelligent Resource Allocation in the Internet of Everything

ResearchDGX agent

arXiv:2607.26602v1 Announce Type: cross Abstract: The rapid development of the Internet of Everything (IoE) is accelerating the adoption of intelligent applications. However, the massive number of con

Low-Precision Training of Large Language Models: Methods, Challenges, and Opportunities

ResearchDGX agent

arXiv:2505.01043v2 Announce Type: replace Abstract: Large language models (LLMs) have achieved impressive performance across various domains. However, the substantial hardware resources required for t

RL^2-VLA: Adaptive RL Latent Compositional Steering with Test-Time Scaling for Vision-Language-Action Models

SafetyDGX agent

arXiv:2607.26991v1 Announce Type: new Abstract: Despite the impressive visuomotor capabilities enabled by Vision-Language-Action (VLA) models, their performance often degrades on challenging and out-o

SPROUT: A Scalable Diffusion Foundation Model for Agricultural Vision

ResearchDGX agent

arXiv:2603.27519v2 Announce Type: replace Abstract: Image-based plant phenotyping depends on dense structural understanding of crops, yet pixel-level annotation remains expensive across species, organ

Using large language models to probe the limits of atom-centered structural descriptors

ResearchDGX agent

arXiv:2607.26984v1 Announce Type: cross Abstract: Mapping an atomic structure to a compact set of geometric descriptors is an essential step in any machine-learning application to atomic-scale modelin

Where Physics Meets Privacy: Federated PINNs for Privacy-Preserving Brain Tumor Biomechanical Modeling

Local AiDGX agent

arXiv:2607.26207v1 Announce Type: new Abstract: Brain tumors such as glioma, meningioma, and pituitary adenoma alter the mechanical behavior of soft brain tissue, yet common diagnostic methods rely on

29 Jul 2026

AGMark: Attention-Guided Dynamic Watermarking for Large Vision-Language Models

ResearchDGX agent

arXiv:2602.09611v2 Announce Type: replace-cross Abstract: Watermarking has emerged as a pivotal solution for content traceability and intellectual property protection in large vision language models (

Configuring Dedicated Model Inference

ToolsDGX agent

The Together AI platform’s dedicated inference architecture consists of three immutable entities: **configs** (engine, GPU type/count, parallelism and optimization profile), **deployments** (a specifi

Explicit Layer Modeling for Video Object Insertion and Layer Decomposition

TutorialsDGX agent

arXiv:2607.25802v1 Announce Type: new Abstract: Most video editing systems still lack explicit layered video representations, limiting their ability to perform realistic compositing, object reuse, and

Exploring Line Bundle Standard Models with Transformers

ResearchDGX agent

arXiv:2607.00078v2 Announce Type: cross Abstract: We propose a Transformer-based Reinforcement Learning architecture, 'LB-Explorer', to search for heterotic line bundle standard models arising from co

Reading Without a Reader: Large Language Models Collapse Reading and Writing into a Single Entangled Code

ApplicationsDGX agent

arXiv:2607.24797v1 Announce Type: cross Abstract: In the literate human brain, reading and writing are two doubly-dissociable systems: a ventral decoding route (impaired in pure alexia) and a fronto-p

SepPrune:A Separator-based Pruning Framework for Efficient Multimodal Large Language Models

ResearchDGX agent

arXiv:2607.25818v1 Announce Type: new Abstract: Recent multimodal large language models (MLLMs), such as Qwen2.5-VL and InternVL3, generate large numbers of vision tokens for high-resolution inputs, l

Specula: Scaling formal specifications for autonomous model checking of system code

AgentsDGX agent

arXiv:2607.25333v1 Announce Type: cross Abstract: Specula is a push-button agentic system that generates high-quality formal specifications for large, complex system code and uses the specifications f

28 Jul 2026

An adaptive multi-fuzzy logic model for diagnosing transformer faults using dynamic weight optimization

ResearchDGX agent

arXiv:2607.23486v1 Announce Type: new Abstract: Dissolved gas analysis (DGA) is crucial for diagnosing early power transformer failures. Traditional DGA interpretation methods like Duval Triangle, IEC

Attention-Guided Layer Selection for Contrastive Decoding in Large Language Models

ResearchDGX agent

arXiv:2607.23067v1 Announce Type: cross Abstract: Contrastive decoding methods such as DoLa improve the factuality of Large Language Models (LLMs) by contrasting the output distributions of mature and

Comparing Optimization Models for Radiotherapy Scheduling

ResearchDGX agent

arXiv:2607.22539v1 Announce Type: cross Abstract: The Radiotherapy Scheduling Problem (RTSP) involves determining an optimal schedule for patients undergoing radiation treatments, a task that has a ma

← Previous
1…202203204205206…1017
Next →