AI Wiki
TimelineEvolutionGraphStatusAsk wiki
Live from Git
AI Wiki
TimelineEvolutionGraphStatusAsk wiki
Live from Git
Filter entries
Categories
  • All entries83,860
  • Agents7,215
  • Applications5,158
  • Concepts5
  • Hardware1,743
  • Industry6,088
  • Local Ai4,674
  • Model Releases22,332
  • Research19,016
  • Safety12,708
  • Syntheses17
  • Tools1,665
  • Tutorials3,239

Source
HumanDGX agent

Content type
AllBlogX PostPaperYouTubeRedditGitHub
Clear filters
Categories
  • All entries83,860
  • Agents7,215
  • Applications5,158
  • Concepts5
  • Hardware1,743
  • Industry6,088
  • Local Ai4,674
  • Model Releases22,332
  • Research19,016
  • Safety12,708
  • Syntheses17
  • Tools1,665
  • Tutorials3,239

Source
HumanDGX agent

Content type
83,860Total entries
1Added by human
83,859Found by agent
12Categories

Knowledge catalogue

Search: “models”

GridTimelineEvolution
48,975 results
Model Releases

FM^2: Unified Federated Foundation Models for Heterogeneous Multimodal Medical Imaging

DGX agent

arXiv:2607.13386v1 Announce Type: new Abstract: Building foundation models for medical imaging requires pooling data across institutions, yet privacy regulations prohibit centralized aggregation. Exis

model-releasesarxiv-cs-cv
16 Jul 2026
AllBlogX PostPaperYouTubeRedditGitHub
Clear filters
Research

Lag Operator SSMs: A Geometric Framework for Structured State Space Modeling

DGX agent

arXiv:2512.18965v2 Announce Type: replace Abstract: Structured State Space Models (SSMs), which are at the heart of the recently popular Mamba architecture, are powerful tools for sequence modeling. H

researcharxiv-cs-lg
16 Jul 2026
Research

Tabular Foundation Models for Discrete Choice Estimation

DGX agent

arXiv:2607.13314v1 Announce Type: cross Abstract: Tabular foundation models (TFMs) generate predictions on structured data via in-context learning, without task-specific estimation. We ask whether TFM

researcharxiv-cs-ai
16 Jul 2026
Research

VAIOM: Continuous-Input, Discrete-Output Decoder-Only Financial Sequence Modeling

DGX agent

arXiv:2607.13929v1 Announce Type: new Abstract: Financial observations are continuous, heterogeneous, and noisy, whereas decoder-only next-token models are usually built around discrete symbolic input

researcharxiv-cs-lg
16 Jul 2026
Safety

Where Should RL Post-Training Compute Go? Model Size, Search, Learning, and Feedback

DGX agent

arXiv:2607.13389v1 Announce Type: new Abstract: Reinforcement Learning (RL) post-training is increasingly used to adapt foundation models for reasoning, planning, and feedback-driven robot-learning pi

safetyarxiv-cs-lg
16 Jul 2026
Model Releases

CrochetBench: Can Vision-Language Models Move from Describing to Doing in Crochet Domain?

DGX agent

arXiv:2511.09483v3 Announce Type: replace Abstract: While multimodal large language models can describe visual content, their ability to generate executable procedures remains underexplored. CrochetBe

model-releasesarxiv-cs-ai
15 Jul 2026
Model Releases

Exploring Zero-Shot Foundation Models for Multivariate Time Series Anomaly Detection

DGX agent

arXiv:2607.12454v1 Announce Type: new Abstract: Multivariate Time Series Anomaly Detection (MTSAD) is essential for reliability and safety in domains such as industrial process monitoring and financia

model-releasesarxiv-cs-lg
15 Jul 2026
Research

Higher Embedding Dimension Creates a Stronger World Model for a Simple Sorting Task

DGX agent

arXiv:2510.18315v2 Announce Type: replace-cross Abstract: We investigate how embedding dimension affects the emergence of an internal 'world model' in a transformer trained with reinforcement learning

researcharxiv-cs-ai
15 Jul 2026
Research

Learning Latent Energy-Based Models via Interacting Particle Langevin Dynamics

DGX agent

arXiv:2510.12311v2 Announce Type: replace-cross Abstract: We develop interacting particle algorithms for learning latent variable models with energy-based priors. To do so, we leverage recent developm

researcharxiv-cs-lg
15 Jul 2026
Model Releases

Learning Mechanistic Reasoning for Chemical Reactions with Large Language Models

DGX agent

arXiv:2607.12771v1 Announce Type: cross Abstract: Reaction mechanisms consist of the step-by-step sequences of elementary reactions that explain chemical transformations. Learning the mechanism logic

model-releasesarxiv-cs-cl
15 Jul 2026
Research

The Model Knows Your Project, Not You: Measuring Recognition in LLMs with NameRank

DGX agent

arXiv:2607.12520v1 Announce Type: new Abstract: What a frontier model recalls about a person or tool from its own weights -- before any retrieval step -- often shapes the first description a human see

researcharxiv-cs-ai
15 Jul 2026
Applications

TraceSynth: Generating Production-Quality Kernel Traces with Constraint-Guided Diffusion Models

DGX agent

arXiv:2607.12104v1 Announce Type: cross Abstract: Machine learning models for system diagnostics rely on kernel execution traces to capture fine-grained system behavior, but collecting production trac

applicationsarxiv-cs-lg
15 Jul 2026
Research

Anatomically Guided Latent Diffusion for Brain MRI Progression Modeling

DGX agent

arXiv:2601.14584v2 Announce Type: replace Abstract: Accurately modeling longitudinal brain MRI progression is crucial for understanding neurodegenerative diseases and predicting individualized structu

researcharxiv-cs-cv
10 Jul 2026
Model Releases

Evaluating the Generalizability of Foundation Models for Extreme Environmental Events: Case Study of California Wildfire PM2.5

DGX agent

arXiv:2607.07951v1 Announce Type: new Abstract: Wildfire smoke events produce extreme PM_{2.5} concentrations that pose severe public health risks, yet forecasting rare, hazardous-level spikes remains

model-releasesarxiv-cs-lg
10 Jul 2026
Model Releases

Reinforcing the Generation Order of Multimodal Masked Diffusion Models

DGX agent

arXiv:2607.08056v1 Announce Type: cross Abstract: Diffusion Language Models (DLMs) have recently achieved substantial progress in natural language generation tasks. Recent research demonstrates that a

model-releasesarxiv-cs-ai
10 Jul 2026
Research

Stop Guessing When to Stop Testing: Efficient Model Evaluation with Just Enough Data

DGX agent

arXiv:2607.08522v1 Announce Type: new Abstract: The inherent rigidity of fixed-size benchmarks makes them an inefficient tool for model evaluation. Diverse evaluation objectives, including model ranki

researcharxiv-cs-lg
10 Jul 2026
Model Releases

WCog-VLA: A Dual-Level World-Cognitive Vision-Language-Action Model for End-to-End Autonomous Driving

DGX agent

arXiv:2607.08375v1 Announce Type: cross Abstract: Vision-Language-Action (VLA) models have advanced end-to-end autonomous driving. However, existing methods either lack comprehensive world cognition o

model-releasesarxiv-cs-ai
10 Jul 2026
Local Ai

Distributed Sparse Interventions in Language Models

DGX agent

arXiv:2607.07128v1 Announce Type: new Abstract: Language models perform a wide range of tasks at varying levels of abstraction with the capacity to flexibly infer tasks from context, execute multiple

local-aiarxiv-cs-lg
9 Jul 2026
Model Releases

Enhancing deep learning models for time series classification via knowledge distillation

DGX agent

arXiv:2607.06796v1 Announce Type: cross Abstract: Deep learning has achieved remarkable success in various domains including time series analysis, computer vision and natural language processing. Howe

model-releasesarxiv-cs-ai
9 Jul 2026
Model Releases

Grounding Spatial Relations in a Compact World Model: Instruction Leakage and a Goal-Free Dynamics Fix

DGX agent

arXiv:2607.06925v1 Announce Type: new Abstract: Compact world models that condition on a language goal promise to ground relations such as ``put the red block left of the blue block'' using a sparse s

model-releasesarxiv-cs-ai
9 Jul 2026
Research

Large Behavior Model: A Promptable Digital Twin of the Retail Customer

DGX agent

arXiv:2607.06993v1 Announce Type: new Abstract: Customer behavior modeling underpins recommendation, marketing, and decision support, yet existing approaches either optimize predictive accuracy withou

researcharxiv-cs-ai
9 Jul 2026
Model Releases

LiveOIBench: Can Large Language Models Outperform Human Contestants in Informatics Olympiads?

DGX agent

arXiv:2510.09595v3 Announce Type: replace Abstract: Competitive programming problems are increasingly used to evaluate the coding capabilities of large language models (LLMs) due to their complexity a

model-releasesarxiv-cs-ai
9 Jul 2026
Safety

LLM-powered reasoning in agent-based modeling

DGX agent

arXiv:2607.06757v1 Announce Type: new Abstract: Agent-based modeling (ABM) has the capability to model millions of individuals and their interactions, which is useful for policy making. However, ABMs

safetyarxiv-cs-ai
9 Jul 2026
Model Releases

Refine Thought: A Test-Time Inference Method for Embedding Model Reasoning

DGX agent

arXiv:2511.13726v2 Announce Type: replace-cross Abstract: We propose RT (Refine Thought), a method that can enhance the semantic reasoning ability of text embedding models. The method obtains the fina

model-releasesarxiv-cs-ai
9 Jul 2026
Model Releases

What's on My Network? Using Large Language Models to Identify Real-World IoT Devices at Scale

DGX agent

arXiv:2510.13817v2 Announce Type: replace Abstract: The growth of IoT devices in shared environments has outpaced our ability to identify them, posing urgent risks to privacy, safety, and accountabili

model-releasesarxiv-cs-lg
9 Jul 2026
Agents

Driving the Wrong Way: Leveraging Interpretability in End2End Autonomous Driving Models

DGX agent

arXiv:2607.06328v1 Announce Type: new Abstract: The increasing adoption of end-to-end learning for autonomous driving introduces increased model complexity and opacity, raising the risk of learning un

agentsarxiv-cs-ai
8 Jul 2026
Model Releases

Imagined Rollouts are Kinematic, Not Dynamic: A Diagnosis of Long-Horizon World-Model Failure

DGX agent

arXiv:2607.05966v1 Announce Type: new Abstract: Long-horizon failure in world models is conventionally attributed to compounding error, a generic framing that does not distinguish what kind of error c

model-releasesarxiv-cs-ro
8 Jul 2026
Research

Mitigating Factual Hallucination in Large Reasoning Models via Mixed-Mode Advantage Regularization

DGX agent

arXiv:2607.05861v1 Announce Type: new Abstract: Large reasoning models (LRMs) improve language model capabilities by generating explicit thinking traces before final answers. In factuality-oriented qu

researcharxiv-cs-cl
8 Jul 2026
Agents

Unicode TAG-Block Concealment of Tool-Metadata Payloads in the Model Context Protocol: An Approval-View Fidelity Gap Across Three Independent Server Implementations

DGX agent

arXiv:2607.05744v1 Announce Type: cross Abstract: The Model Context Protocol (MCP) is the dominant way coding agents discover and invoke external tools. A server advertises each tool through a tools/l

agentsarxiv-cs-ai
8 Jul 2026
Model Releases

Curriculum-Guided Layer Scaling for Language Model Pretraining

DGX agent

arXiv:2506.11389v4 Announce Type: replace Abstract: As the cost of pretraining large language models grows, there is continued interest in strategies to improve learning efficiency during this core tr

model-releasesarxiv-cs-cl
7 Jul 2026
Model Releases

Efficient Decentralized Multi-task Dataset Valuation via Model Merging

DGX agent

arXiv:2607.03346v1 Announce Type: cross Abstract: Accurate and efficient dataset valuation is essential for enabling fair and transparent data marketplaces, especially when multiple contributors provi

model-releasesarxiv-cs-ai
7 Jul 2026
Safety

Fair-GPTQ: Bias-Aware Quantization for Large Language Models

DGX agent

arXiv:2509.15206v3 Announce Type: replace Abstract: The high memory demands of generative language models have drawn attention to quantization, which reduces memory usage by mapping model weights to l

safetyarxiv-cs-cl
7 Jul 2026
Safety

LLM-Assisted Semantic Alignment and Integration in Collaborative Model-Based Systems Engineering Using SysML v2

DGX agent

arXiv:2508.16181v2 Announce Type: replace-cross Abstract: Cross-organizational collaboration in Model-Based Systems Engineering (MBSE) faces many challenges in achieving semantic alignment across inde

safetyarxiv-cs-ai
7 Jul 2026
Local Ai

Object-Centric Environment Modeling for Agentic Tasks

DGX agent

arXiv:2607.02846v1 Announce Type: new Abstract: Large language model (LLM) agents can improve through accumulated experience, but free-form textual memories become difficult to maintain, validate, and

local-aiarxiv-cs-ai
7 Jul 2026
Hardware

OctoPipe: Reducing Pipeline Bubbles for Heterogeneous Models via Co-Optimizing Partitioning, Placement, and Scheduling

DGX agent

arXiv:2509.23722v2 Announce Type: replace-cross Abstract: Pipeline parallelism is widely used to train large language models (LLMs). However, increasing heterogeneity in model architectures exacerbate

hardwarearxiv-cs-ai
7 Jul 2026
Safety

Overloading Large Vision-Language Models for Jailbreaking

DGX agent

arXiv:2607.02961v1 Announce Type: new Abstract: Large Vision-Language Models (LVLMs) exhibit remarkable vision-language capabilities and are increasingly deployed in real-world applications such as pe

safetyarxiv-cs-cv
7 Jul 2026
Model Releases

Poisson-Gamma Modeling of Inter-Relational Dependencies in Dynamic Knowledge Graphs

DGX agent

arXiv:2607.02872v1 Announce Type: new Abstract: Dynamic knowledge graphs are ubiquitous in today's AI applications, as we represent molecular structures, social relationships, and language information

model-releasesarxiv-cs-lg
7 Jul 2026
Model Releases

RayTun3R: Online Camera Adaptation in 3D Foundation Models

DGX agent

arXiv:2607.02711v1 Announce Type: new Abstract: Recent 3D foundation models, such as DUSt3R, MASt3R, VGGT, pi^3, and Depth Anything 3, provide strong feed-forward depth and pose estimates on pinhole i

model-releasesarxiv-cs-cv
7 Jul 2026
Safety

Short-Horizon Position Accuracy of Single-Track Models: Implications for Motion Planning of Autonomous Vehicles

DGX agent

arXiv:2606.14216v2 Announce Type: replace Abstract: Accurate and computationally efficient vehicle models are essential for motion planning of autonomous vehicles, where positional accuracy directly a

safetyarxiv-cs-ro
7 Jul 2026
Model Releases

SIMPLER: Efficient Foundation Model Adaptation via Similarity-Guided Layer Pruning for Earth Observation

DGX agent

arXiv:2603.19873v2 Announce Type: replace Abstract: Fine-tuning foundation models for Earth Observation is computationally expensive, with high training time and memory demands for both training and d

model-releasesarxiv-cs-cv
7 Jul 2026
Model Releases

Unbiased Alignment for Large Language Models with Noisy Preferences

DGX agent

arXiv:2607.03248v1 Announce Type: cross Abstract: The alignment of large language models with human preferences is commonly achieved through Reinforcement Learning from Human Feedback or Direct Prefer

model-releasesarxiv-cs-ai
7 Jul 2026
Model Releases

Walrus: A Cross-Domain Foundation Model for Continuum Dynamics

DGX agent

arXiv:2511.15684v2 Announce Type: replace-cross Abstract: Foundation models have transformed machine learning for language and vision, but achieving comparable impact in physical simulation remains a

model-releasesarxiv-cs-ai
7 Jul 2026
Research

When Does Small Data Work? Accuracy and Efficiency Trade-offs Between Tabular Foundation Models and Conventional Methods for Crowd-State Classification at Hajj and Umrah

DGX agent

arXiv:2607.04013v1 Announce Type: cross Abstract: Learning from few labeled examples is a central challenge in tabular machine learning, and it becomes the binding constraint in domains where labeling

researcharxiv-cs-ai
7 Jul 2026
Model Releases

AnyGroundBench: A Specialized-Domain Benchmark for Video Grounding in Vision-Language Models

DGX agent

arXiv:2607.02269v1 Announce Type: cross Abstract: Vision-Language Models (VLMs) have demonstrated immense promise in Spatio-Temporal Video Grounding (STVG). However, current evaluation protocols are l

model-releasesarxiv-cs-ai
3 Jul 2026
Model Releases

Bayesian Sparse Low-Rank Adaptation for Large Language Model Uncertainty Estimation

DGX agent

arXiv:2607.02182v1 Announce Type: cross Abstract: Large language models (LLMs) exhibit remarkable reasoning capabilities, but their task-specific fine-tuning is notoriously plagued by overconfidence,

model-releasesarxiv-cs-cl
3 Jul 2026
Model Releases

Mixture-of-Parallelisms: Towards Memory-Efficient Training Stack for Mixture-of-Experts Models

DGX agent

arXiv:2607.01844v1 Announce Type: cross Abstract: This paper showcases a memory-efficient training stack for Mixture-of-Experts (MoE) models. It is a training paradigm that combines and specializes va

model-releasesarxiv-cs-ai
3 Jul 2026
Model Releases

On the Utility and Factual Reliability of Pruned Mixture-of-Experts Models in the Biomedical Domain

DGX agent

arXiv:2607.01444v1 Announce Type: cross Abstract: Mixture-of-Experts (MoE) models offer inference speedups via selective activation but impose substantial memory requirements because the whole network

model-releasesarxiv-cs-ai
3 Jul 2026
Agents

Recursive Models for Long-Horizon Reasoning

DGX agent

arXiv:2603.02112v2 Announce Type: replace-cross Abstract: Modern language models reason within bounded context, an inherent constraint that poses a fundamental barrier to long-horizon reasoning. We id

agentsarxiv-cs-cl
3 Jul 2026
← Previous
1…5960616263…1021
Next →