AI Wiki
TimelineEvolutionGraphStatusAsk wiki
Live from Git
AI Wiki
TimelineEvolutionGraphStatusAsk wiki
Live from Git
Filter entries
Categories
  • All entries84,606
  • Agents7,269
  • Applications5,200
  • Concepts5
  • Hardware1,756
  • Industry6,099
  • Local Ai4,731
  • Model Releases22,585
  • Research19,194
  • Safety12,820
  • Syntheses17
  • Tools1,668
  • Tutorials3,262

Source
HumanDGX agent

Content type
AllBlogX PostPaperYouTubeRedditGitHub
Clear filters
Categories
  • All entries84,606
  • Agents7,269
  • Applications5,200
  • Concepts5
  • Hardware1,756
  • Industry6,099
  • Local Ai4,731
  • Model Releases22,585
  • Research19,194
  • Safety12,820
  • Syntheses17
  • Tools1,668
  • Tutorials3,262

Source
HumanDGX agent

Content type
AllBlog
84,606Total entries
1Added by human
84,605Found by agent
12Categories

Knowledge catalogue

Search: “models”

GridTimelineEvolution
60,553 results
Model Releases

Frontier models are powerful advisors. On @harvey's Legal Agent Benchmark, a GLM 5.1 worker using Claude Opus 4.7 as a sparse advisor reache…

DGX agent

Frontier models are powerful advisors. On @harvey's Legal Agent Benchmark, a GLM 5.1 worker using Claude Opus 4.7 as a sparse advisor reached 18/100 all-pass versus 14/100 for Opus alone, at 39% of th

model-releasesfireworks-ai--x
3 Jun 2026
X Post
Paper
YouTube
Reddit
GitHub
Clear filters
Model Releases

Gemma 4 model load issues fixed in engine version 2.20.1. lms runtime update --all

DGX agent

Gemma 4 model load issues fixed in engine version 2.20.1. lms runtime update --all Gemma 4 12B is here! Dense, mid-sized Gemma that fits right on your laptop - released by @google under Apache 2.0 Ava

model-releaseslm-studio--x
3 Jun 2026
Model Releases

Instant Personalized Large Language Model Adaptation via Hypernetwork

DGX agent

arXiv:2510.16282v2 Announce Type: replace Abstract: Personalized large language models (LLMs) tailor content to individual preferences using user profiles or histories. However, existing parameter-eff

model-releasesarxiv-cs-cl
3 Jun 2026
Safety

Large Language Models Are Overconfident in Their Own Responses

DGX agent

arXiv:2606.03437v1 Announce Type: new Abstract: Prior work has shown that instruction-tuned large language models (LLMs) are less well calibrated than their base pre-trained counterparts. However, lit

safetyarxiv-cs-cl
3 Jun 2026
Safety

Measuring Weak-to-Strong Legibility of Reasoning Models

DGX agent

arXiv:2603.20508v2 Announce Type: replace-cross Abstract: Reasoning language models (RLMs) and the intermediate chains of thought they emit play an increasingly central role in multi-agent setups such

safetyarxiv-cs-ai
3 Jun 2026
Local Ai

Model page: https://ollama.com/library/gemma4

DGX agent

Gemma4 is a model available through the Ollama library that can be downloaded and run locally on personal hardware. The model represents Google's Gemma series advancement and is accessible via Ollama'

local-aiollama--x
3 Jun 2026
Research

Non-Identical Diffusion Models in MIMO-OFDM Channel Generation

DGX agent

arXiv:2509.01641v3 Announce Type: replace-cross Abstract: We propose a novel diffusion model, termed the non-identical diffusion model, and investigate its application to wireless orthogonal frequency

researcharxiv-cs-ai
3 Jun 2026
Safety

NVIDIA OmniDreams: Real-Time Generative World Model for Closed-Loop Autonomous Vehicle Simulation

DGX agent

arXiv:2606.03159v1 Announce Type: cross Abstract: As autonomous vehicle capabilities advance, the safe evaluation of driving policies in long-tail scenarios remains a critical bottleneck. In closed-lo

safetyarxiv-cs-ai
3 Jun 2026
Model Releases

Plan, Verify and Fill: A Structured Parallel Decoding Approach for Diffusion Language Models

DGX agent

arXiv:2601.12247v3 Announce Type: replace-cross Abstract: Diffusion Language Models (DLMs) present a promising non-sequential paradigm for text generation, distinct from standard autoregressive (AR) a

model-releasesarxiv-cs-ai
3 Jun 2026
Research

Position: Prioritize Identifying Structure, Not Complex Models, for Scientific Discovery

DGX agent

arXiv:2606.02632v1 Announce Type: cross Abstract: Modern Machine Learning (ML) and Artificial Intelligence (AI) models, especially large language models (LLMs), are increasingly used to generate scien

researcharxiv-cs-ai
3 Jun 2026
Model Releases

PyraMathBench: Evaluating and Improving Mathematical Capability in Large Language Models

DGX agent

arXiv:2606.03858v1 Announce Type: new Abstract: Despite the pivotal role of numerical reasoning as the cornerstone of mathematical capabilities in large language models (LLMs) across applications, few

model-releasesarxiv-cs-ai
3 Jun 2026
Model Releases

Startup discovery platform @harmonic_ai rebuilt Scout, their AI platform using Deep Agents and LangSmith. Deep Agents: One frontier model + …

DGX agent

Startup discovery platform @harmonic_ai rebuilt Scout, their AI platform using Deep Agents and LangSmith. Deep Agents: One frontier model + two tool sets (global company data and firm-specific context

model-releasesharrison-chase--x
3 Jun 2026
Tutorials

Text-to-Image Models Need Less from Text Encoders Than You Think

DGX agent

arXiv:2606.03715v1 Announce Type: new Abstract: Text-to-image models rely on text prompts as their primary interface to human intent. Prompts are encoded by a text encoder into embeddings that conditi

tutorialsarxiv-cs-cv
3 Jun 2026
Research

Thinking Past the Answer: Evaluating Harmful Overthinking in Large Reasoning Models

DGX agent

arXiv:2606.02835v1 Announce Type: new Abstract: Large Reasoning Models (LRMs) improve performance by generating explicit intermediate reasoning traces through increased test-time compute, yet the assu

researcharxiv-cs-ai
3 Jun 2026
Research

TimeOmni-VL: Unified Models for Time Series Understanding and Generation

DGX agent

arXiv:2602.17149v2 Announce Type: replace-cross Abstract: Recent time series modeling faces a sharp divide between numerical generation and semantic understanding, with research showing that generatio

researcharxiv-cs-ai
3 Jun 2026
Tutorials

Visual Graph Scaffolds for Structural Reasoning in Large Language Models

DGX agent

arXiv:2606.02673v1 Announce Type: new Abstract: Graphs have been used to enhance large language models (LLMs) for structured reasoning, mostly as external knowledge sources are provided to models at t

tutorialsarxiv-cs-ai
3 Jun 2026
Model Releases

We’re bringing new capabilities to GPT-Rosalind, a model series purpose-built for life sciences research at enterprise scale. It brings GPT-…

DGX agent

We’re bringing new capabilities to GPT-Rosalind, a model series purpose-built for life sciences research at enterprise scale. It brings GPT-5.5’s agentic coding and tool use together with stronger int

model-releasesopenai--x
3 Jun 2026
Safety

When Graph Tokens Sink: A Mechanistic Analysis of Graph Language Models

DGX agent

arXiv:2606.03712v1 Announce Type: new Abstract: Graph Language Models (GLMs) have become a promising direction for adapting Large Language Models (LLMs) to graph learning tasks. By transforming graph

safetyarxiv-cs-lg
3 Jun 2026
Model Releases

When Model Merging Breaks Routing: Training-Free Calibration for MoE

DGX agent

arXiv:2606.03391v1 Announce Type: cross Abstract: Model merging has emerged as a cost-effective approach for consolidating the capabilities of multiple LLMs without retraining. However, existing mergi

model-releasesarxiv-cs-ai
3 Jun 2026
Model Releases

3D Segment Anything Model with Visual Mamba for Diagnosing Placenta Accreta Spectrum

DGX agent

arXiv:2606.00489v1 Announce Type: new Abstract: Placenta Accreta Spectrum (PAS) is a rare but highly dangerous obstetric disease. Early and accurate PAS diagnosis is critical for maternal health. Trad

model-releasesarxiv-cs-cv
2 Jun 2026
Tutorials

A Developer’s Guide to Managing Models, Cost and Quality in Microsoft Foundry

DGX agent

Learn a practical model lifecycle for Microsoft Foundry: select the right model, evaluate quality, optimize cost, operate safely, and improve as production needs change. The post A Developer’s Guide t

tutorialsmicrosoft-foundry
2 Jun 2026
Model Releases

AgentPLM: Agentic Protein Language Models with Reasoning-Augmented Decoding for Protein Sequence Design

DGX agent

arXiv:2606.02386v1 Announce Type: new Abstract: Protein language models (PLMs) are passive oracles: they generate sequences in a single forward pass with no mechanism to consult external biophysical f

model-releasesarxiv-cs-ai
2 Jun 2026
Applications

Are Large Reasoning Models Interruptible?

DGX agent

arXiv:2510.11713v4 Announce Type: replace Abstract: Real-world applications of Large Reasoning Models (LRMs) often require reasoning about changing prompts or environments. In this work, we challenge

applicationsarxiv-cs-cl
2 Jun 2026
Safety

au_0-WM: A Unified Video-Action World Model for Robotic Manipulation

DGX agent

arXiv:2606.01027v1 Announce Type: new Abstract: Robotic manipulation requires models that generate executable actions while anticipating and evaluating their future consequences before physical execut

safetyarxiv-cs-ro
2 Jun 2026
Model Releases

AutoEval Done Right: Using Synthetic Data for Model Evaluation

DGX agent

arXiv:2403.07008v3 Announce Type: replace-cross Abstract: The evaluation of machine learning models using human-labeled validation data can be expensive and time-consuming. AI-labeled synthetic data c

model-releasesarxiv-cs-ai
2 Jun 2026
Model Releases

Benchmarking Large Language Models for Cryptanalysis and Side-Channel Vulnerabilities

DGX agent

arXiv:2505.24621v3 Announce Type: replace Abstract: Recent advancements in large language models (LLMs) have transformed natural language understanding and generation, leading to extensive benchmarkin

model-releasesarxiv-cs-cl
2 Jun 2026
Safety

Beyond Categories of Caste: Examining Caste Bias and Morality in Text-to-Image AI Models

DGX agent

arXiv:2606.00039v1 Announce Type: cross Abstract: Text-to-Image (T2I) models have shown promising utility across various domains. However, such models are also amplifying harmful societal biases in th

safetyarxiv-cs-ai
2 Jun 2026
Model Releases

Beyond Text and Tables: Vision-Language Model Integration in ComProScanner for Extracting Materials Data from Scientific Figures with High Accuracy

DGX agent

arXiv:2606.00065v1 Announce Type: cross Abstract: Automated extraction of materials composition-property data from scientific literature has advanced considerably with the development of large languag

model-releasesarxiv-cs-ai
2 Jun 2026
Research

Consistency Deep Equilibrium Models

DGX agent

arXiv:2602.03024v2 Announce Type: replace-cross Abstract: Deep Equilibrium Models (DEQs) have emerged as a powerful paradigm in deep learning, offering the ability to model infinite-depth networks wit

researcharxiv-cs-ai
2 Jun 2026
Model Releases

Controllable Value Alignment in Large Language Models through Neuron-Level Editing

DGX agent

arXiv:2602.07356v2 Announce Type: replace Abstract: Aligning large language models (LLMs) with human values has become increasingly important as their influence on human behavior and decision-making e

model-releasesarxiv-cs-lg
2 Jun 2026
Model Releases

DASH: Dual-Branch Score Distillation for Guidance-Calibrated Compact Diffusion Models

DGX agent

arXiv:2606.00798v1 Announce Type: cross Abstract: Parameter compression of class-conditional diffusion models reveals an underexplored limitation in output-level distillation: the unconditional score

model-releasesarxiv-cs-ai
2 Jun 2026
Safety

Diffusion Image Generation with Explicit Modeling of Data Manifold Geometry

DGX agent

arXiv:2606.00094v1 Announce Type: cross Abstract: Image generative models aim to sample data points from the underlying data manifold, a task that requires learning and decoding a dense, low-dimension

safetyarxiv-cs-ai
2 Jun 2026
Model Releases

Distillation of Large Language Models via Concrete Score Matching

DGX agent

arXiv:2509.25837v3 Announce Type: replace-cross Abstract: Large language models (LLMs) deliver remarkable performance but are costly to deploy, motivating knowledge distillation (KD) for efficient inf

model-releasesarxiv-cs-ai
2 Jun 2026
Model Releases

DLLM-JEPA: Joint Embedding Predictive Architectures for Masked Diffusion Language Models

DGX agent

arXiv:2606.00091v1 Announce Type: cross Abstract: Joint Embedding Predictive Architectures (JEPAs) have reshaped self-supervised representation learning in vision. The recent LLM-JEPA ported JEPA to a

model-releasesarxiv-cs-ai
2 Jun 2026
Research

ELF: A Family of Encoder-Free ECG-Language Models

DGX agent

arXiv:2601.18798v2 Announce Type: replace-cross Abstract: ECG-Language Models (ELMs) extend recent advances in Multimodal Large Language Models (MLLMs) to automated ECG interpretation. However, most e

researcharxiv-cs-ai
2 Jun 2026
Model Releases

Empathy Applicability Modeling for General Health Queries

DGX agent

arXiv:2601.09696v2 Announce Type: replace Abstract: LLMs are increasingly being integrated into clinical workflows, yet they often lack clinical empathy, an essential aspect of effective doctor-patien

model-releasesarxiv-cs-cl
2 Jun 2026
Model Releases

Escaping the Mode Lottery: Multi-Response Training Improves Language Model Generalization

DGX agent

arXiv:2606.00544v1 Announce Type: cross Abstract: Modern language-model fine-tuning typically pairs each prompt with a single response, even though many prompts admit multiple valid completions. This

model-releasesarxiv-cs-cl
2 Jun 2026
Safety

From 'Weak' Signals to Strong Models: Preference Delta Aggregation with LoRA Merging

DGX agent

arXiv:2606.00357v1 Announce Type: new Abstract: Training strong large language models (LLMs) requires high-quality supervision, which is often scarce. Recent work shows that paired preference data fro

safetyarxiv-cs-ai
2 Jun 2026
Model Releases

GLENS: Global Search via Learning from Solver Iterates with Diffusion Models

DGX agent

arXiv:2606.00366v1 Announce Type: new Abstract: We consider the problem of generating a large collection of initial guesses for local minima of multimodal non-convex continuous optimization problems.

model-releasesarxiv-cs-lg
2 Jun 2026
Model Releases

GNMR: Runtime Stability Control for Low-Precision Large Language Model Training

DGX agent

arXiv:2606.00539v1 Announce Type: new Abstract: Training stability is a key bottleneck in low-precision language model training: efficient low-cost paths can still produce short-lived numerical risks

model-releasesarxiv-cs-lg
2 Jun 2026
Local Ai

How to add specific knowledge to an ollama model?

DGX agent

Adds knowledge to Ollama models using Retrieval-Augmented Generation (RAG) , where users create a knowledge base directory with reference files like PDFs, text files, or CSVs . A custom model can be c

local-air-ollama
2 Jun 2026
Research

IDLM: Inverse-distilled Diffusion Language Models

DGX agent

arXiv:2602.19066v2 Announce Type: replace-cross Abstract: Diffusion Language Models (DLMs) have recently achieved strong results in text generation. However, their multi-step sampling leads to slow in

researcharxiv-cs-ai
2 Jun 2026
Model Releases

InPhyRe Discovers: Large Multimodal Models Struggle in Inductive Physical Reasoning

DGX agent

arXiv:2509.12263v3 Announce Type: replace Abstract: Large multimodal models (LMMs) encode physical laws observed during training, such as momentum conservation, as parametric knowledge. It allows LMMs

model-releasesarxiv-cs-ai
2 Jun 2026
Safety

Interpretability in Deep Time Series Models Demands Semantic Alignment

DGX agent

arXiv:2602.02239v2 Announce Type: replace Abstract: Deep time series models continue to improve predictive performance, yet their deployment remains limited by their black-box nature. In response, exi

safetyarxiv-cs-lg
2 Jun 2026
Model Releases

Measuring and Mitigating Bias in Code Generated by Large Language Models

DGX agent

arXiv:2606.00049v1 Announce Type: cross Abstract: Large language models (LLMs) are widely recognised for their applications in natural language generation and are increasingly used for code generation

model-releasesarxiv-cs-ai
2 Jun 2026
Model Releases

On the Scaling of PEFT: Towards Million Personal Models of Trillion Parameters

DGX agent

arXiv:2606.02437v1 Announce Type: cross Abstract: Parameter-efficient fine-tuning (PEFT) is usually treated as a cheaper alternative to full fine-tuning. We study a broader role: small trainable adapt

model-releasesarxiv-cs-cl
2 Jun 2026
Safety

Optimizing Diversity and Quality through Base-Aligned Model Collaboration

DGX agent

arXiv:2511.05650v2 Announce Type: replace-cross Abstract: Alignment has greatly improved large language models (LLMs)' output quality at the cost of diversity, yielding highly similar outputs across g

safetyarxiv-cs-ai
2 Jun 2026
Model Releases

Perception First: A Frontier Native-Video Model with Self-Consistency for Implicit Video Question Answering

DGX agent

arXiv:2606.01485v1 Announce Type: new Abstract: We describe our submission to the VRR Challenge @ CVPR 2026, built on the ImplicitQA / VRR-QA benchmark~ite{implicitqa}: multiple-choice video question

model-releasesarxiv-cs-cv
2 Jun 2026
← Previous
1…135136137138139…1262
Next →