AI Wiki
TimelineEvolutionGraphStatusAsk wiki
Live from Git
AI Wiki
TimelineEvolutionGraphStatusAsk wiki
Live from Git
Filter entries
Categories
  • All entries83,860
  • Agents7,215
  • Applications5,158
  • Concepts5
  • Hardware1,743
  • Industry6,088
  • Local Ai4,674
  • Model Releases22,332
  • Research19,016
  • Safety12,708
  • Syntheses17
  • Tools1,665
  • Tutorials3,239

Source
HumanDGX agent

Content type
AllBlogX PostPaperYouTubeRedditGitHub
Clear filters
Categories
  • All entries83,860
  • Agents7,215
  • Applications5,158
  • Concepts5
  • Hardware1,743
  • Industry6,088
  • Local Ai4,674
  • Model Releases22,332
  • Research19,016
  • Safety12,708
  • Syntheses17
  • Tools1,665
  • Tutorials3,239

Source
HumanDGX agent

83,860Total entries
1Added by human
83,859Found by agent
12Categories

Knowledge catalogue

Search: “models”

GridTimelineEvolution
59,929 results
16 Jul 2026

Design, Modeling and Experimental Validation of a Miniature Hybrid Underwater Glider With Large-Range Foldable Deflectable Wings

Model ReleasesDGX agent

arXiv:2607.13622v1 Announce Type: new Abstract: Miniature hybrid underwater gliders have attracted increasing attention for long-endurance ocean observation and confined-space inspection. Large-range

FM^2: Unified Federated Foundation Models for Heterogeneous Multimodal Medical Imaging

Model ReleasesDGX agent

arXiv:2607.13386v1 Announce Type: new Abstract: Building foundation models for medical imaging requires pooling data across institutions, yet privacy regulations prohibit centralized aggregation. Exis

Lag Operator SSMs: A Geometric Framework for Structured State Space Modeling

ResearchDGX agent
Content type
AllBlogX PostPaperYouTubeRedditGitHub
Clear filters

arXiv:2512.18965v2 Announce Type: replace Abstract: Structured State Space Models (SSMs), which are at the heart of the recently popular Mamba architecture, are powerful tools for sequence modeling. H

Open models are already being used in the enterprise. Over 85% of the Fortune 500 companies already use Ollama to fulfill specific tasks. @j…

Local AiDGX agent

Open models are already being used in the enterprise. Over 85% of the Fortune 500 companies already use Ollama to fulfill specific tasks. @jmorgan Why open models and Ollama? Ownership. Open models ar

Tabular Foundation Models for Discrete Choice Estimation

ResearchDGX agent

arXiv:2607.13314v1 Announce Type: cross Abstract: Tabular foundation models (TFMs) generate predictions on structured data via in-context learning, without task-specific estimation. We ask whether TFM

VAIOM: Continuous-Input, Discrete-Output Decoder-Only Financial Sequence Modeling

ResearchDGX agent

arXiv:2607.13929v1 Announce Type: new Abstract: Financial observations are continuous, heterogeneous, and noisy, whereas decoder-only next-token models are usually built around discrete symbolic input

Where Should RL Post-Training Compute Go? Model Size, Search, Learning, and Feedback

SafetyDGX agent

arXiv:2607.13389v1 Announce Type: new Abstract: Reinforcement Learning (RL) post-training is increasingly used to adapt foundation models for reasoning, planning, and feedback-driven robot-learning pi

15 Jul 2026

CrochetBench: Can Vision-Language Models Move from Describing to Doing in Crochet Domain?

Model ReleasesDGX agent

arXiv:2511.09483v3 Announce Type: replace Abstract: While multimodal large language models can describe visual content, their ability to generate executable procedures remains underexplored. CrochetBe

Exploring Zero-Shot Foundation Models for Multivariate Time Series Anomaly Detection

Model ReleasesDGX agent

arXiv:2607.12454v1 Announce Type: new Abstract: Multivariate Time Series Anomaly Detection (MTSAD) is essential for reliability and safety in domains such as industrial process monitoring and financia

Generic AI models are not built for medicine. Doximity Ask, trained and served on Fireworks, outperformed GPT-5.6 Sol, Claude Fable 5, and O…

Model ReleasesDGX agent

Generic AI models are not built for medicine. Doximity Ask, trained and served on Fireworks, outperformed GPT-5.6 Sol, Claude Fable 5, and OpenEvidence in a Stanford-Harvard clinical AI safety study.

Higher Embedding Dimension Creates a Stronger World Model for a Simple Sorting Task

ResearchDGX agent

arXiv:2510.18315v2 Announce Type: replace-cross Abstract: We investigate how embedding dimension affects the emergence of an internal 'world model' in a transformer trained with reinforcement learning

Interesting surprise drop from Thinky! The Inkling model looks pretty solid on benchmarks, and it has some little surprises in its architect…

Model ReleasesDGX agent

Interesting surprise drop from Thinky! The Inkling model looks pretty solid on benchmarks, and it has some little surprises in its architecture: - Small conv layers in several places - An RMSNorm for

Learning Latent Energy-Based Models via Interacting Particle Langevin Dynamics

ResearchDGX agent

arXiv:2510.12311v2 Announce Type: replace-cross Abstract: We develop interacting particle algorithms for learning latent variable models with energy-based priors. To do so, we leverage recent developm

Learning Mechanistic Reasoning for Chemical Reactions with Large Language Models

Model ReleasesDGX agent

arXiv:2607.12771v1 Announce Type: cross Abstract: Reaction mechanisms consist of the step-by-step sequences of elementary reactions that explain chemical transformations. Learning the mechanism logic

The Model Knows Your Project, Not You: Measuring Recognition in LLMs with NameRank

ResearchDGX agent

arXiv:2607.12520v1 Announce Type: new Abstract: What a frontier model recalls about a person or tool from its own weights -- before any retrieval step -- often shapes the first description a human see

TraceSynth: Generating Production-Quality Kernel Traces with Constraint-Guided Diffusion Models

ApplicationsDGX agent

arXiv:2607.12104v1 Announce Type: cross Abstract: Machine learning models for system diagnostics rely on kernel execution traces to capture fine-grained system behavior, but collecting production trac

10 Jul 2026

Anatomically Guided Latent Diffusion for Brain MRI Progression Modeling

ResearchDGX agent

arXiv:2601.14584v2 Announce Type: replace Abstract: Accurately modeling longitudinal brain MRI progression is crucial for understanding neurodegenerative diseases and predicting individualized structu

Evaluating the Generalizability of Foundation Models for Extreme Environmental Events: Case Study of California Wildfire PM2.5

Model ReleasesDGX agent

arXiv:2607.07951v1 Announce Type: new Abstract: Wildfire smoke events produce extreme PM_{2.5} concentrations that pose severe public health risks, yet forecasting rare, hazardous-level spikes remains

Not everyone will succeed in the model business. 'You've seen a lot of companies spend a huge amount on compute and then not end up with a g…

Model ReleasesDGX agent

Cohere warns that computational investment alone does not guarantee success in building AI models, emphasizing that many companies spend substantial resources on computing infrastructure without achie

Reinforcing the Generation Order of Multimodal Masked Diffusion Models

Model ReleasesDGX agent

arXiv:2607.08056v1 Announce Type: cross Abstract: Diffusion Language Models (DLMs) have recently achieved substantial progress in natural language generation tasks. Recent research demonstrates that a

Stop Guessing When to Stop Testing: Efficient Model Evaluation with Just Enough Data

ResearchDGX agent

arXiv:2607.08522v1 Announce Type: new Abstract: The inherent rigidity of fixed-size benchmarks makes them an inefficient tool for model evaluation. Diverse evaluation objectives, including model ranki

WCog-VLA: A Dual-Level World-Cognitive Vision-Language-Action Model for End-to-End Autonomous Driving

Model ReleasesDGX agent

arXiv:2607.08375v1 Announce Type: cross Abstract: Vision-Language-Action (VLA) models have advanced end-to-end autonomous driving. However, existing methods either lack comprehensive world cognition o

9 Jul 2026

Distributed Sparse Interventions in Language Models

Local AiDGX agent

arXiv:2607.07128v1 Announce Type: new Abstract: Language models perform a wide range of tasks at varying levels of abstraction with the capacity to flexibly infer tasks from context, execute multiple

Enhancing deep learning models for time series classification via knowledge distillation

Model ReleasesDGX agent

arXiv:2607.06796v1 Announce Type: cross Abstract: Deep learning has achieved remarkable success in various domains including time series analysis, computer vision and natural language processing. Howe

Grounding Spatial Relations in a Compact World Model: Instruction Leakage and a Goal-Free Dynamics Fix

Model ReleasesDGX agent

arXiv:2607.06925v1 Announce Type: new Abstract: Compact world models that condition on a language goal promise to ground relations such as ``put the red block left of the blue block'' using a sparse s

Is Meta AI back? Haven't seen Mark post in three years here. Plus, the model is available via API. Not to mention the courage to announce it…

Model ReleasesDGX agent

Is Meta AI back? Haven't seen Mark post in three years here. Plus, the model is available via API. Not to mention the courage to announce it the same week as the long-awaited GPT-5.6. Great timing if

Large Behavior Model: A Promptable Digital Twin of the Retail Customer

ResearchDGX agent

arXiv:2607.06993v1 Announce Type: new Abstract: Customer behavior modeling underpins recommendation, marketing, and decision support, yet existing approaches either optimize predictive accuracy withou

LiveOIBench: Can Large Language Models Outperform Human Contestants in Informatics Olympiads?

Model ReleasesDGX agent

arXiv:2510.09595v3 Announce Type: replace Abstract: Competitive programming problems are increasingly used to evaluate the coding capabilities of large language models (LLMs) due to their complexity a

LLM-powered reasoning in agent-based modeling

SafetyDGX agent

arXiv:2607.06757v1 Announce Type: new Abstract: Agent-based modeling (ABM) has the capability to model millions of individuals and their interactions, which is useful for policy making. However, ABMs

Meta launches flagship Muse Spark 1.1 model with multi-agent upgrades

Model ReleasesDGX agent

Meta Platforms Inc. today launched a new flagship large language model optimized to power multi-agent automation workflows. Muse Spark 1.1 is available in the company’s Meta AI chatbot service and via

Refine Thought: A Test-Time Inference Method for Embedding Model Reasoning

Model ReleasesDGX agent

arXiv:2511.13726v2 Announce Type: replace-cross Abstract: We propose RT (Refine Thought), a method that can enhance the semantic reasoning ability of text embedding models. The method obtains the fina

What's on My Network? Using Large Language Models to Identify Real-World IoT Devices at Scale

Model ReleasesDGX agent

arXiv:2510.13817v2 Announce Type: replace Abstract: The growth of IoT devices in shared environments has outpaced our ability to identify them, posing urgent risks to privacy, safety, and accountabili

8 Jul 2026

Driving the Wrong Way: Leveraging Interpretability in End2End Autonomous Driving Models

AgentsDGX agent

arXiv:2607.06328v1 Announce Type: new Abstract: The increasing adoption of end-to-end learning for autonomous driving introduces increased model complexity and opacity, raising the risk of learning un

Imagined Rollouts are Kinematic, Not Dynamic: A Diagnosis of Long-Horizon World-Model Failure

Model ReleasesDGX agent

arXiv:2607.05966v1 Announce Type: new Abstract: Long-horizon failure in world models is conventionally attributed to compounding error, a generic framing that does not distinguish what kind of error c

Meta launches image generation model with coding, search capabilities

Model ReleasesDGX agent

Meta Platforms Inc. today debuted an image generation model that can write code and search the web. Muse Image is the second algorithm released to date by Meta Superintelligence Labs, the company’s ar

Mitigating Factual Hallucination in Large Reasoning Models via Mixed-Mode Advantage Regularization

ResearchDGX agent

arXiv:2607.05861v1 Announce Type: new Abstract: Large reasoning models (LRMs) improve language model capabilities by generating explicit thinking traces before final answers. In factuality-oriented qu

Unicode TAG-Block Concealment of Tool-Metadata Payloads in the Model Context Protocol: An Approval-View Fidelity Gap Across Three Independent Server Implementations

AgentsDGX agent

arXiv:2607.05744v1 Announce Type: cross Abstract: The Model Context Protocol (MCP) is the dominant way coding agents discover and invoke external tools. A server advertises each tool through a tools/l

We've usually stayed away from model comparisons but 5.6 vs Fable is a unique situation We've never had a case where the team is so complete…

Model ReleasesDGX agent

We've usually stayed away from model comparisons but 5.6 vs Fable is a unique situation We've never had a case where the team is so completely convinced on which one is better Here's the timeline of o

7 Jul 2026

Curriculum-Guided Layer Scaling for Language Model Pretraining

Model ReleasesDGX agent

arXiv:2506.11389v4 Announce Type: replace Abstract: As the cost of pretraining large language models grows, there is continued interest in strategies to improve learning efficiency during this core tr

Efficient Decentralized Multi-task Dataset Valuation via Model Merging

Model ReleasesDGX agent

arXiv:2607.03346v1 Announce Type: cross Abstract: Accurate and efficient dataset valuation is essential for enabling fair and transparent data marketplaces, especially when multiple contributors provi

Fair-GPTQ: Bias-Aware Quantization for Large Language Models

SafetyDGX agent

arXiv:2509.15206v3 Announce Type: replace Abstract: The high memory demands of generative language models have drawn attention to quantization, which reduces memory usage by mapping model weights to l

LLM-Assisted Semantic Alignment and Integration in Collaborative Model-Based Systems Engineering Using SysML v2

SafetyDGX agent

arXiv:2508.16181v2 Announce Type: replace-cross Abstract: Cross-organizational collaboration in Model-Based Systems Engineering (MBSE) faces many challenges in achieving semantic alignment across inde

Object-Centric Environment Modeling for Agentic Tasks

Local AiDGX agent

arXiv:2607.02846v1 Announce Type: new Abstract: Large language model (LLM) agents can improve through accumulated experience, but free-form textual memories become difficult to maintain, validate, and

OctoPipe: Reducing Pipeline Bubbles for Heterogeneous Models via Co-Optimizing Partitioning, Placement, and Scheduling

HardwareDGX agent

arXiv:2509.23722v2 Announce Type: replace-cross Abstract: Pipeline parallelism is widely used to train large language models (LLMs). However, increasing heterogeneity in model architectures exacerbate

Overloading Large Vision-Language Models for Jailbreaking

SafetyDGX agent

arXiv:2607.02961v1 Announce Type: new Abstract: Large Vision-Language Models (LVLMs) exhibit remarkable vision-language capabilities and are increasingly deployed in real-world applications such as pe

Poisson-Gamma Modeling of Inter-Relational Dependencies in Dynamic Knowledge Graphs

Model ReleasesDGX agent

arXiv:2607.02872v1 Announce Type: new Abstract: Dynamic knowledge graphs are ubiquitous in today's AI applications, as we represent molecular structures, social relationships, and language information

RayTun3R: Online Camera Adaptation in 3D Foundation Models

Model ReleasesDGX agent

arXiv:2607.02711v1 Announce Type: new Abstract: Recent 3D foundation models, such as DUSt3R, MASt3R, VGGT, pi^3, and Depth Anything 3, provide strong feed-forward depth and pose estimates on pinhole i

Short-Horizon Position Accuracy of Single-Track Models: Implications for Motion Planning of Autonomous Vehicles

SafetyDGX agent

arXiv:2606.14216v2 Announce Type: replace Abstract: Accurate and computationally efficient vehicle models are essential for motion planning of autonomous vehicles, where positional accuracy directly a

SIMPLER: Efficient Foundation Model Adaptation via Similarity-Guided Layer Pruning for Earth Observation

Model ReleasesDGX agent

arXiv:2603.19873v2 Announce Type: replace Abstract: Fine-tuning foundation models for Earth Observation is computationally expensive, with high training time and memory demands for both training and d

Unbiased Alignment for Large Language Models with Noisy Preferences

Model ReleasesDGX agent

arXiv:2607.03248v1 Announce Type: cross Abstract: The alignment of large language models with human preferences is commonly achieved through Reinforcement Learning from Human Feedback or Direct Prefer

Walrus: A Cross-Domain Foundation Model for Continuum Dynamics

Model ReleasesDGX agent

arXiv:2511.15684v2 Announce Type: replace-cross Abstract: Foundation models have transformed machine learning for language and vision, but achieving comparable impact in physical simulation remains a

When Does Small Data Work? Accuracy and Efficiency Trade-offs Between Tabular Foundation Models and Conventional Methods for Crowd-State Classification at Hajj and Umrah

ResearchDGX agent

arXiv:2607.04013v1 Announce Type: cross Abstract: Learning from few labeled examples is a central challenge in tabular machine learning, and it becomes the binding constraint in domains where labeling

6 Jul 2026

Most models run on whatever inference engine they ship with. That's rarely the fastest option. We built our own, AGI-RUN. On a phone it beat…

Local AiDGX agent

Most models run on whatever inference engine they ship with. That's rarely the fastest option. We built our own, AGI-RUN. On a phone it beat every model's own native engine, and ran up to 9.9x faster

3 Jul 2026

AnyGroundBench: A Specialized-Domain Benchmark for Video Grounding in Vision-Language Models

Model ReleasesDGX agent

arXiv:2607.02269v1 Announce Type: cross Abstract: Vision-Language Models (VLMs) have demonstrated immense promise in Spatio-Temporal Video Grounding (STVG). However, current evaluation protocols are l

Bayesian Sparse Low-Rank Adaptation for Large Language Model Uncertainty Estimation

Model ReleasesDGX agent

arXiv:2607.02182v1 Announce Type: cross Abstract: Large language models (LLMs) exhibit remarkable reasoning capabilities, but their task-specific fine-tuning is notoriously plagued by overconfidence,

Mixture-of-Parallelisms: Towards Memory-Efficient Training Stack for Mixture-of-Experts Models

Model ReleasesDGX agent

arXiv:2607.01844v1 Announce Type: cross Abstract: This paper showcases a memory-efficient training stack for Mixture-of-Experts (MoE) models. It is a training paradigm that combines and specializes va

On the Utility and Factual Reliability of Pruned Mixture-of-Experts Models in the Biomedical Domain

Model ReleasesDGX agent

arXiv:2607.01444v1 Announce Type: cross Abstract: Mixture-of-Experts (MoE) models offer inference speedups via selective activation but impose substantial memory requirements because the whole network

Recursive Models for Long-Horizon Reasoning

AgentsDGX agent

arXiv:2603.02112v2 Announce Type: replace-cross Abstract: Modern language models reason within bounded context, an inherent constraint that poses a fundamental barrier to long-horizon reasoning. We id

Restoring Linguistic Grounding in VLA Models via Train-Free Attention Recalibration

Model ReleasesDGX agent

arXiv:2603.06001v2 Announce Type: replace-cross Abstract: Vision-Language-Action (VLA) models enable robots to perform manipulation tasks directly from natural language instructions and are increasing

Unlocking Speech-Text Compositional Powers: Instruction-Following Speech Language Models without Instruction Tuning

ResearchDGX agent

arXiv:2607.02214v1 Announce Type: new Abstract: Instruction tuning for speech language models (SLMs) is substantially more challenging than for text-based large language models (LLMs), as it requires

← Previous
1…6162636465…999
Next →