AI Wiki
TimelineEvolutionGraphStatusAsk wiki
Live from Git
AI Wiki
TimelineEvolutionGraphStatusAsk wiki
Live from Git
Filter entries
Categories
  • All entries84,588
  • Agents7,266
  • Applications5,200
  • Concepts5
  • Hardware1,756
  • Industry6,098
  • Local Ai4,730
  • Model Releases22,577
  • Research19,194
  • Safety12,816
  • Syntheses17
  • Tools1,667
  • Tutorials3,262

Source
HumanDGX agent

Content type
AllBlogX PostPaperYouTubeRedditGitHub
Clear filters
Categories
  • All entries84,588
  • Agents7,266
  • Applications5,200
  • Concepts5
  • Hardware1,756
  • Industry6,098
  • Local Ai4,730
  • Model Releases22,577
  • Research19,194
  • Safety12,816
  • Syntheses17
  • Tools1,667
  • Tutorials3,262

Source
HumanDGX agent

Content type
84,588Total entries
1Added by human
84,587Found by agent
12Categories

Knowledge catalogue

Search: “models”

GridTimelineEvolution
60,536 results
Model Releases

What’s new in AI infrastructure and orchestration this month

DGX agent

At Google, AI is a soup-to-nuts endeavor. Obviously, we make leading AI models like Gemini and Nano Banana. We incorporate AI into the tools you use every day (think Gmail, BigQuery, AlloyDB, Google C

model-releasesgoogle-cloud-ai
31 Jul 2026
AllBlogX PostPaperYouTubeRedditGitHub
Clear filters
Model Releases

Why are AI model tests always the same generic prompts?

DGX agent

Okay, hear me out. Why is it that every time a new model comes out, all the tests I see are 'make a car game,' 'make a website,' or something equally generic, usually from a prompt that's barely a lin

model-releasesr-localllama
31 Jul 2026
Safety

Anatomy Contextualized Adaption of CT Foundation Models

DGX agent

arXiv:2607.27154v1 Announce Type: new Abstract: CT vision-language foundation models have demonstrated promising performance across downstream tasks, but are typically trained with whole-volume repres

safetyarxiv-cs-cv
30 Jul 2026
Safety

Dual Inversion for Text-to-Image Diffusion Models: From Both Prompt and Noise Perspectives

DGX agent

arXiv:2607.26735v1 Announce Type: new Abstract: Prompt inversion, as a typical reverse engineering technique, enables text-to-image (T2I) diffusion models to generate the desired target images without

safetyarxiv-cs-cv
30 Jul 2026
Safety

From Found to Designed: Concepts as a Design Axis for Large Language Models

DGX agent

arXiv:2607.26825v1 Announce Type: new Abstract: Large language models (LLMs) encode rich concept-like information, but represent it implicitly through distributed statistical associations rather than

safetyarxiv-cs-cl
30 Jul 2026
Model Releases

Rad-JEPA 3D: Radiology Joint-Embedding Predictive Model for 3D Computed Tomography

DGX agent

arXiv:2607.26196v1 Announce Type: new Abstract: Self-supervised pretraining is central to 3D medical image analysis, where unlabeled CT volumes are abundant but expert annotations are scarce. Yet exis

model-releasesarxiv-cs-cv
30 Jul 2026
Applications

See2Think: Do Multimodal Models Really Use Intermediate Visual States?

DGX agent

arXiv:2607.26769v1 Announce Type: new Abstract: Multimodal large language models increasingly use sketches, annotations, tools, and intermediate images during reasoning, but it remains unclear whether

applicationsarxiv-cs-cv
30 Jul 2026
Local Ai

Try Again, Don't Look Back: Blind Resampling Outperforms Self-Repair in Small Code Models

DGX agent

arXiv:2607.26117v1 Announce Type: cross Abstract: Self-repair - returning a failed program to the model together with its test output and asking for a correction - is a standard component of code agen

local-aiarxiv-cs-lg
30 Jul 2026
Model Releases

Two Calls Beat Five Agents: Evaluating Multi-Agent Pipelines Against Self-Refinement for Local Language Models

DGX agent

arXiv:2607.26922v1 Announce Type: new Abstract: Multi-agent LLM pipeline systems break down the task among multiple roles for better reasoning, but are benchmarked mainly with large-scale commercial m

model-releasesarxiv-cs-lg
30 Jul 2026
Model Releases

What Can Latent World Models Know? Physical Parameter Identifiability in Multimodal Predictive Representations

DGX agent

arXiv:2607.27017v1 Announce Type: new Abstract: A central premise of latent world models is that predicting the future forces a representation to internalize the physics of its environment. Which phys

model-releasesarxiv-cs-lg
30 Jul 2026
Research

Accurate structural modeling of chemically diverse molecular interfaces with Vilya-2

DGX agent

arXiv:2607.25156v1 Announce Type: new Abstract: Structure-prediction networks built on co-evolutionary statistics have transformed protein-based drug discovery, yet their accuracy does not extend to p

researcharxiv-cs-lg
29 Jul 2026
Model Releases

BREAKING: SpaceXAI's newly released Grok Voice Think Fast 2.0 beats voice models from OpenAI, Google, Alibaba, and DeepSlate in the Artifici…

DGX agent

SpaceXAI has released its new Grok Voice Think Fast 2.0, which on the Artificial Analysis Speech‑to‑Speech benchmark outperformed leading models from OpenAI, Google, Alibaba and DeepSlate. The claim w

model-releaseselon-musk--x
29 Jul 2026
Research

CaRE Compute-aware Remasking Evaluation Protocol for Masked Diffusion Language Models

DGX agent

arXiv:2607.24763v1 Announce Type: new Abstract: Masked diffusion language models (MDLMs) are advancing rapidly, yet the evaluation standards needed to reliably interpret their progress have not kept p

researcharxiv-cs-ai
29 Jul 2026
Applications

Empirical Evaluation of Out-Of-Distribution Performance of Tabular Foundation Models

DGX agent

arXiv:2607.26000v1 Announce Type: cross Abstract: Tabular Foundation Models (TFMs) have emerged as novel approaches for tabular predictive tasks, demonstrating competitive predictive performance to en

applicationsarxiv-cs-ai
29 Jul 2026
Model Releases

GrocLM: Grocery Category Recommendation in E-Commerce with Large Language Models

DGX agent

arXiv:2607.24764v1 Announce Type: new Abstract: The rapid growth of online grocery shopping requires recommendation systems that capture cyclical purchasing behavior and diverse user intents. Traditio

model-releasesarxiv-cs-ai
29 Jul 2026
Safety

Med-SegLens: Latent-Level Model Diffing for Interpretable Medical Image Segmentation

DGX agent

arXiv:2602.10508v2 Announce Type: replace Abstract: Modern segmentation models achieve strong predictive performance but remain largely opaque, limiting our ability to diagnose failures, understand da

safetyarxiv-cs-cv
29 Jul 2026
Applications

Physics of Language Models: Part 4.1, Architecture Design and the Magic of Canon Layers

DGX agent

arXiv:2512.17351v2 Announce Type: replace Abstract: Understanding architectural differences in language models is challenging, especially at academic-scale pretraining (e.g., 1.3B parameters, 100B tok

applicationsarxiv-cs-cl
29 Jul 2026
Local Ai

SAM3D-Guided Object-Centric Representation Alignment for Vision-Language-Action Models

DGX agent

arXiv:2607.25912v1 Announce Type: cross Abstract: Vision-Language-Action (VLA) models have shown strong potential for general robot manipulation, but most existing models rely on 2D visual-language ba

local-aiarxiv-cs-ai
29 Jul 2026
Model Releases

Sharpness-aware Model Merging with Salience Recovery for LLM-based Cross-Domain Sequential Recommendation

DGX agent

arXiv:2607.25366v1 Announce Type: cross Abstract: LLM-based Cross-Domain Sequential Recommendation (CDSR) leverages LLMs to enhance target performance via deep semantic reasoning, alleviating the depe

model-releasesarxiv-cs-lg
29 Jul 2026
Tutorials

VisualPatchWorld: Code World Models as Latent Structured Representations for Planning

DGX agent

arXiv:2607.25236v1 Announce Type: new Abstract: Different research lines use the term world model in different ways, yet they share a common aim: to capture how the world evolves under action in a for

tutorialsarxiv-cs-cl
29 Jul 2026
Model Releases

We’re giving scientists, mathematicians, and engineers free access to our frontier models—starting with 10,000 researchers and expanding to …

DGX agent

We’re giving scientists, mathematicians, and engineers free access to our frontier models—starting with 10,000 researchers and expanding to 100,000 through 2027. ChatGPT for Academic Researchers is bu

model-releasesopenai--x
29 Jul 2026
Model Releases

Agenta: an open-source Claude Cowork alternative where you can use self-hosted models (and any harness)

DGX agent

Hey r/LocalLLaMA, I’m Mahmoud from Agenta. We built a self-hosted, more flexible, alternative to Claude Cowork . This short video shows how it works. I use it to build AI coworkers for my startup, lik

model-releasesr-localllama
28 Jul 2026
Industry

AI model compression startup Multiverse raises 570M at 1.7B valuation

DGX agent

Multiverse Computing SL, a startup working on technology that compresses artificial intelligence models so that they run more efficiently on less hardware, announced Monday it raised 570 million in Se

industrysiliconangle
28 Jul 2026
Hardware

Anthropic faces backlash from Silicon Valley partners, founders, and researchers for competitive tactics, guardrails, and lack of support for open-weight models (Wall Street Journal)

DGX agent

Wall Street Journal: Anthropic faces backlash from Silicon Valley partners, founders, and researchers for competitive tactics, guardrails, and lack of support for open-weight models — The AI pioneer f

hardwaretechmeme
28 Jul 2026
Research

{au}: Learning Touch-Augmented Vision-Language-Action Models from Future Visual Supervision

DGX agent

arXiv:2607.24485v1 Announce Type: new Abstract: Learning the informative tactile representation while effectively adapting it to pretrained Vision-Language-Action (VLA) models remains challenging at b

researcharxiv-cs-ro
28 Jul 2026
Model Releases

Code Arena now measures fullstack capabilities! View overall rankings across AI models on full-stack web development tasks: multi-step reaso…

DGX agent

Code Arena now measures fullstack capabilities! View overall rankings across AI models on full-stack web development tasks: multi-step reasoning, tool use, and end-to-end app generation. - Kimi K3 (Ma

model-releaseskimi-moonshot--x
28 Jul 2026
Research

Continuous surrogates versus threshold Boolean networks for modeling Arabidopsis ISR gene regulation

DGX agent

arXiv:2607.23289v1 Announce Type: cross Abstract: Gene regulatory network modeling often requires balancing predictive accuracy and mechanistic interpretability. In this work, we compare continuous su

researcharxiv-cs-lg
28 Jul 2026
Research

DINOv3-MIL: Per-Kidney Multi-Label Tumour and Cyst Detection from Foundation-Model Patch Tokens on KiTS23

DGX agent

arXiv:2607.22687v1 Announce Type: new Abstract: Foundation vision models trained on natural images transfer to medical tasks without domain pre-training, but volumetric classification requires aggrega

researcharxiv-cs-cv
28 Jul 2026
Model Releases

Do Diagrams Help Large Language Models Reason? Evidence from Syllogistic Reasoning

DGX agent

arXiv:2607.23513v1 Announce Type: cross Abstract: Diagrams are widely used to support logical reasoning, and prior studies suggest that representations such as Euler diagrams can improve human reasoni

model-releasesarxiv-cs-ai
28 Jul 2026
Model Releases

Domain-Prior-Regularized Graph Modeling for Anomaly Detection in Cyber-Physical Systems

DGX agent

arXiv:2607.23197v1 Announce Type: new Abstract: Anomaly detection on multivariate sensor time series is critical for industrial monitoring of cyber-physical systems (CPS), where even subtle deviations

model-releasesarxiv-cs-lg
28 Jul 2026
Model Releases

Extending Fourier Neural Operators for Modeling Parameterized and Coupled PDEs

DGX agent

arXiv:2607.23466v1 Announce Type: new Abstract: Parameterized and coupled partial differential equations (PDEs) are central to modeling phenomena in science and engineering, yet neural operator method

model-releasesarxiv-cs-lg
28 Jul 2026
Research

FeelWorld: Visuo-Tactile World Model for Hierarchical Contact Prediction and Planning

DGX agent

arXiv:2607.24267v1 Announce Type: new Abstract: Humans plan physical interactions by imagining the possible outcomes of candidate actions. However, existing visual world models primarily capture appea

researcharxiv-cs-ro
28 Jul 2026
Research

Gradient Networks for Universal Magnetic Modeling of Synchronous Machines

DGX agent

arXiv:2602.14947v2 Announce Type: replace-cross Abstract: This paper presents a physics-constrained neural network framework for dynamic modeling of saturable synchronous machines, including spatial h

researcharxiv-cs-lg
28 Jul 2026
Research

Masked Distillation: Internalizing the Chain-of-Thought in Language Models

DGX agent

arXiv:2607.22629v1 Announce Type: new Abstract: Large Reasoning Models (LRMs) produce long, explicit chains of intermediate steps before generating a final answer at inference time. These intermediate

researcharxiv-cs-ai
28 Jul 2026
Model Releases

MEGA-CL: A Molecular Foundation Model for Generalizable ADMET Prediction through Graph External Attention and Contrastive Learning

DGX agent

arXiv:2607.24314v1 Announce Type: new Abstract: Predicting the absorption, distribution, metabolism, excretion and toxicity (ADMET) properties of small molecules remains a major challenge in drug disc

model-releasesarxiv-cs-lg
28 Jul 2026
Research

Metamorphic Testing for Clinical ML Models: A Framework Proposal and Pilot Study

DGX agent

arXiv:2607.22984v1 Announce Type: cross Abstract: Machine learning models for clinical prediction tasks, such as in-hospital mortality and sepsis onset, routinely achieve high AUROC scores. However, A

researcharxiv-cs-lg
28 Jul 2026
Model Releases

Might need math+code benchmark for frontier model(LLMs Silently Replace Math)[D]

DGX agent

Hello guys. I found some problems in current frontier models. And want to share. # math_code_hallucination > Record of a failure caused by combining mathematics and code in a single prompt. --- ## Cas

model-releasesr-machinelearning
28 Jul 2026
Model Releases

MixQuant: Adaptive Mixed-Precision Quantization for Large Language Models

DGX agent

arXiv:2607.23047v1 Announce Type: cross Abstract: Mixed-precision quantization improves the accuracy of post-training quantization by allocating higher bitwidths to sensitive layers, but existing meth

model-releasesarxiv-cs-ai
28 Jul 2026
Agents

Multi-agent DRL-based Lane Change Decision Model for Cooperative Platooning in Mixed Traffic

DGX agent

arXiv:2601.11809v2 Announce Type: replace Abstract: Connected automated vehicles (CAVs) possess the ability to communicate and coordinate with one another, enabling cooperative platooning that enhance

agentsarxiv-cs-ai
28 Jul 2026
Research

PCA: Persistence-Aware Compression and Aggregation for Fast Video Large Language Models

DGX agent

arXiv:2607.22726v1 Announce Type: new Abstract: Despite advances in Video Large Language Models (VLLMs) that have displayed promising outcomes in video understanding, the redundancy in the long-durati

researcharxiv-cs-cv
28 Jul 2026
Research

Realizing Scaling Laws in Recommender Systems: A Foundation-Expert Paradigm for Hyperscale Model Deployment

DGX agent

arXiv:2508.02929v3 Announce Type: replace-cross Abstract: Scaling laws have been established for recommender systems, yet efficiently deploying foundation model (FM) across multiple recommendation sur

researcharxiv-cs-ai
28 Jul 2026
Model Releases

Reason Popper-ly: Patching In-Context Reasoning with Inductive Logic Programming

DGX agent

arXiv:2607.23019v1 Announce Type: new Abstract: Chain-of-thought (CoT) prompting enables large language models (LLMs) to tackle multi-step reasoning tasks, yet the generated intermediate steps are not

model-releasesarxiv-cs-ai
28 Jul 2026
Local Ai

Retrieval-Augmented Large Language Models as Components of Cognitive Computing architecture for Regulatory Knowledge Management

DGX agent

arXiv:2607.24352v1 Announce Type: new Abstract: The aim of this article is to verify whether integrating large language models (LLMs) with the Retrieval-Augmented Generation (RAG) architecture enables

local-aiarxiv-cs-cl
28 Jul 2026
Model Releases

SINT-Flow: Schema Integration using Large Language Model Workflows

DGX agent

arXiv:2607.24492v1 Announce Type: new Abstract: The goal of schema integration is, given a set of input schemata or tables, to derive a global, unified schema that is able to represent the concepts, a

model-releasesarxiv-cs-cl
28 Jul 2026
Model Releases

The Few-shot Dilemma: Over-prompting Large Language Models

DGX agent

arXiv:2509.13196v2 Announce Type: replace Abstract: Over-prompting, a phenomenon where excessive examples in prompts lead to diminished performance in Large Language Models (LLMs), challenges the conv

model-releasesarxiv-cs-cl
28 Jul 2026
Model Releases

The Gate Always Closes: On Injecting Auxiliary Signals into Frozen Vision-Language Models

DGX agent

arXiv:2607.23335v1 Announce Type: new Abstract: Auxiliary signal pathways in VLMs are routinely fitted with learnable gates so the optimiser can decide how much of the signal to admit. We find that th

model-releasesarxiv-cs-cv
28 Jul 2026
Model Releases

Through the Bottleneck: How Multi-head Latent Attention Separates Content from Position in Language Models

DGX agent

arXiv:2607.23054v1 Announce Type: cross Abstract: Multi-head Latent Attention (MLA), introduced in DeepSeek-V2, compresses key-value pairs through a shared low-rank bottleneck (cKV), achieving 81% KV-

model-releasesarxiv-cs-ai
28 Jul 2026
Research

Towards Cultural Bridge by Bahnaric-Vietnamese Translation Using Transfer Learning of Sequence-To-Sequence Pre-training Language Model

DGX agent

arXiv:2505.11421v2 Announce Type: replace Abstract: This work explores the journey towards achieving Bahnaric-Vietnamese translation for the sake of culturally bridging the two ethnic groups in Vietna

researcharxiv-cs-cl
28 Jul 2026
← Previous
1…125126127128129…1262
Next →