AI Wiki
TimelineEvolutionGraphStatusAsk wiki
Live from Git
AI Wiki
TimelineEvolutionGraphStatusAsk wiki
Live from Git
Filter entries
Categories
  • All entries83,193
  • Agents7,156
  • Applications5,120
  • Concepts5
  • Hardware1,734
  • Industry6,079
  • Local Ai4,640
  • Model Releases22,098
  • Research18,859
  • Safety12,600
  • Syntheses17
  • Tools1,664
  • Tutorials3,221

Source
HumanDGX agent

Content type
AllBlogX PostPaperYouTubeRedditGitHub
Clear filters
Categories
  • All entries83,193
  • Agents7,156
  • Applications5,120
  • Concepts5
  • Hardware1,734
  • Industry6,079
  • Local Ai4,640
  • Model Releases22,098
  • Research18,859
  • Safety12,600
  • Syntheses17
  • Tools1,664
  • Tutorials3,221

Source
HumanDGX agent

Content type
AllBlog
83,193Total entries
1Added by human
83,192Found by agent
12Categories

Knowledge catalogue

Search: “models”

GridTimelineEvolution
59,399 results
Model Releases

RepBench: Compiling Benchmarks into Capability Representations for Large Language Models

DGX agent

arXiv:2607.28008v1 Announce Type: new Abstract: Representation engineering reads and steers capability directions in large language models, yet methods are typically evaluated on paper-specific synthe

model-releasesarxiv-cs-cl
31 Jul 2026
X Post
Paper
YouTube
Reddit
GitHub
Clear filters
Model Releases

Same Facts, Different Diagnosis: Measuring and Mitigating Narrative Anchoring in Clinical Language Models

DGX agent

arXiv:2607.27384v1 Announce Type: new Abstract: Large language models used for clinical diagnostic reasoning are sensitive to sociolinguistic register, not just clinical content. We term this failure

model-releasesarxiv-cs-cl
31 Jul 2026
Model Releases

Language Models are not Equally Robust to Non-Canonical Tokenization across Languages

DGX agent

arXiv:2607.26831v1 Announce Type: new Abstract: Despite the existence of exponentially many valid tokenizations for a given string, language models operate on a single canonical sequence deterministic

model-releasesarxiv-cs-cl
30 Jul 2026
Model Releases

Laplace-PSN-IRT: Uncertainty Quantification for Neural Item Response Theory Models of LLM Benchmarks

DGX agent

arXiv:2607.25257v1 Announce Type: cross Abstract: Item Response Theory (IRT) has recently been proposed as a framework for evaluating large language model (LLM) benchmarks by separating a model's late

model-releasesarxiv-cs-ai
29 Jul 2026
Model Releases

WALoMA: A Multitask Wireless Foundation Model via Adaptive Low-Rank Masked Autoencoders

DGX agent

arXiv:2607.25763v1 Announce Type: cross Abstract: This paper proposes a multitask wireless foundation model via adaptive low-rank masked autoencoders (WALoMA), a unified multi-task foundation model fo

model-releasesarxiv-cs-lg
29 Jul 2026
Model Releases

AIR-BENCH Live: An Evolving Safety Benchmark for Foundation Models

DGX agent

arXiv:2607.22671v1 Announce Type: new Abstract: Foundation-model safety benchmarks capture the AI risks of their time of publication: as models improve and governments pass new AI-safety legislation,

model-releasesarxiv-cs-ai
28 Jul 2026
Model Releases

Numerical Investigation of Sequence Modeling Theory using Controllable Memory Functions

DGX agent

arXiv:2506.05678v3 Announce Type: replace Abstract: The evolution of sequence modeling architectures, from recurrent neural networks and convolutional models to Transformers and structured state-space

model-releasesarxiv-cs-lg
28 Jul 2026
Research

Reverso: Efficient Time Series Foundation Models for Zero-shot Forecasting

DGX agent

arXiv:2602.17634v2 Announce Type: replace-cross Abstract: Learning time series foundation models has been shown to be a promising approach for zero-shot time series forecasting across diverse time ser

researcharxiv-cs-ai
28 Jul 2026
Safety

RM-Distiller: Exploiting Generative LLM for Reward Model Distillation

DGX agent

arXiv:2601.14032v2 Announce Type: replace Abstract: Reward models (RMs) play a pivotal role in aligning large language models (LLMs) with human preferences. Due to the difficulty of obtaining high-qua

safetyarxiv-cs-cl
28 Jul 2026
Model Releases

Big update: Among open-weight models, Kimi K3 (Max) is #1 in the Agent Arena with +9.75% net-improvement, surpassing GLM-5.2 (Max) at +7.12%…

DGX agent

Big update: Among open-weight models, Kimi K3 (Max) is #1 in the Agent Arena with +9.75% net-improvement, surpassing GLM-5.2 (Max) at +7.12%, and landed the #1 spot across 5 signals (see below). Kimi

model-releaseskimi-moonshot--x
27 Jul 2026
Model Releases

MoE^2-LoRA: When MoE Models Meet MoE-style Low-Rank Adaptation

DGX agent

arXiv:2607.21978v1 Announce Type: new Abstract: Mixture-of-Experts (MoE) architectures have been widely adopted in large language models, yet parameter-efficient fine-tuning (PEFT) for MoE models rema

model-releasesarxiv-cs-cl
27 Jul 2026
Model Releases

PhantomFill: When the Form Demands an Answer, Language Models Invent One

DGX agent

arXiv:2607.20492v1 Announce Type: cross Abstract: Language models in production do not write prose. They fill forms: JSON fields, function arguments, extraction templates. We show that the form itself

model-releasesarxiv-cs-ai
24 Jul 2026
Model Releases

Task Competence Is Not Instruction Following: Evaluating Instruction-Conflicting Behavior in Small Language Models

DGX agent

arXiv:2607.19608v1 Announce Type: new Abstract: Instruction tuning is meant to make language models follow user requests, yet it is unclear whether small models comply when an instruction conflicts wi

model-releasesarxiv-cs-cl
23 Jul 2026
Model Releases

Toward a Vision-Language Foundation Model for Medical Data: Multimodal Dataset and Benchmarks for Vietnamese PET/CT Report Generation

DGX agent

arXiv:2509.24739v4 Announce Type: replace Abstract: Vision-Language Foundation Models (VLMs), trained on large-scale multimodal datasets, have driven significant advances in Artificial Intelligence (A

model-releasesarxiv-cs-cv
23 Jul 2026
Model Releases

Mixed-Timescale Differential Coding for Downlink Model Broadcast in Wireless Federated Learning

DGX agent

arXiv:2607.13119v1 Announce Type: cross Abstract: In standard federated learning systems, the parameter server broadcasts the global model to the participating devices in every iteration. Motivated by

model-releasesarxiv-cs-lg
16 Jul 2026
Local Ai

TSSM: Triaxial State Space Model for Global Station Weather Forecasting with Temporal-Variable-Historical Modeling

DGX agent

arXiv:2607.13101v1 Announce Type: cross Abstract: Global Station Weather Forecasting (GSWF) is pivotal for localized and extreme weather prediction over key regions. Despite efforts to exploit look-ba

local-aiarxiv-cs-ai
16 Jul 2026
Model Releases

Do We Really Need Multimodal Emotion Language Models Larger Than 1B Parameters?

DGX agent

arXiv:2607.12787v1 Announce Type: new Abstract: Recent advances in multimodal large language models (MLLMs) have significantly improved the performance of multimodal emotion recognition (MER) and enab

model-releasesarxiv-cs-ai
15 Jul 2026
Model Releases

OvisOCR2 (0.8B): first end-to-end model to top OmniDocBench - I threw 827 real scanned medical docs at it, here's everything I learned

DGX agent

What it is: ATH-MaaS/OvisOCR2 - a 0.8B document-parsing VLM post-trained from Qwen3.5-0.8B (SFT + RL + OPD), Apache 2.0, runs on vLLM 0.22.1. One prompt per page image -> complete markdown (HTML table

model-releasesr-localllama
15 Jul 2026
Model Releases

Silent Alarm: A J-Space Protocol for Comparing Danger Recognition Across Models and Quantization Levels

DGX agent

arXiv:2607.12792v1 Announce Type: cross Abstract: Jailbreak-robustness research typically evaluates safety through generated responses using an LLM-as-judge approach. Such evaluations, however, are se

model-releasesarxiv-cs-ai
15 Jul 2026
Model Releases

Google named a Leader in the 2026 IDC MarketScape for Worldwide Foundation Model Software

DGX agent

For years, we’ve built with a clear priority: putting the practical needs of the enterprise first. Long before generative AI dominated the headlines, we were focused on building the global infrastruct

model-releasesgoogle-cloud-ai
14 Jul 2026
Model Releases

New model for AMD Strix Halo users: My 198B Step 3.7 Flash release was a big hit, but this one may be even better: 298B-parameter Hy3, now r…

DGX agent

New model for AMD Strix Halo users: My 198B Step 3.7 Flash release was a big hit, but this one may be even better: 298B-parameter Hy3, now running on a new 2-bit FPX codebook designed to map efficient

model-releasesclem-delangue--x
13 Jul 2026
Local Ai

.@satyanadella's Reverse Information Paradox is real. What @satyanadella's calling for already exists: open models. Over 9M+ developers have…

DGX agent

.@satyanadella's Reverse Information Paradox is real. What @satyanadella's calling for already exists: open models. Over 9M+ developers have used @ollama to access open models and keep their competiti

local-aiollama--x
13 Jul 2026
Model Releases

AUTOPILOT VQA: Benchmarking Vision-Language Models for Incident-Centric Dashcam Understanding

DGX agent

arXiv:2607.08745v1 Announce Type: new Abstract: Recent advances in Vision-Language Models, Large Language Models, and Multimodal Large Language Models have improved autonomous driving tasks such as sc

model-releasesarxiv-cs-ai
10 Jul 2026
Model Releases

Bifidelity Parameter Estimation Using Conditional Diffusion Models

DGX agent

arXiv:2504.01894v2 Announce Type: replace Abstract: We present a bifidelity method for uncertainty quantification of parameter estimates in complex systems, leveraging generative models trained to sam

model-releasesarxiv-cs-lg
9 Jul 2026
Model Releases

GLM 5.2: a new rise of open-weight agentic models

DGX agent

On June 16th, Z.ai released GLM 5.2, its latest flagship model. At the time of announcement, it advertised scores at or near Anthropic and OpenAI's models, and far ahead of GLM 5.1. In the world of us

model-releaseslambda-labs
9 Jul 2026
Model Releases

Deform360: A Massive Multi-view Visuotactile Dataset for Deformable World Models

DGX agent

arXiv:2607.05390v1 Announce Type: cross Abstract: Predicting object dynamics (i.e., world modeling) is a fundamental challenge for robotic manipulation, and modeling deformable objects presents a part

model-releasesarxiv-cs-cv
7 Jul 2026
Local Ai

DELTA-TTS: Adapting Autoregressive Model into Diffusion Language Model for Text-to-Speech

DGX agent

arXiv:2607.04140v1 Announce Type: cross Abstract: Autoregressive (AR) text-to-speech (TTS) models generate discrete speech tokens sequentially, which makes inference slow and can degrade robustness by

local-aiarxiv-cs-cl
7 Jul 2026
Tutorials

Recommended reading if you are scaling with open models. BTW, you should be thinking about how to scale with open-weight models.

DGX agent

This post from DAIR.AI recommends resources for developers and organizations scaling applications using open-weight language models, emphasizing the importance of planning infrastructure and deploymen

tutorialsdair-ai--x
30 Jun 2026
Model Releases

FULL INTERVIEW: Engineers Edward Coristine (@as400495) and Tai Groot (@taigrr) just released an ML model called Rampart for the National Des…

DGX agent

FULL INTERVIEW: Engineers Edward Coristine (@as400495) and Tai Groot (@taigrr) just released an ML model called Rampart for the National Design Studio. It's a local-first, open-source AI privacy model

model-releasesclem-delangue--x
29 Jun 2026
Research

@NousResearch We are working on benching various combos of open source models to see if we can get Opus levels with much cheaper models as w…

DGX agent

Nous Research is conducting benchmarking tests to evaluate whether combinations of open-source models can achieve performance comparable to Anthropic's Claude Opus while maintaining significantly lowe

researchnous-research--x
26 Jun 2026
Safety

Bias Fitting to Mitigate Length Bias of Reward Model in RLHF

DGX agent

arXiv:2505.12843v2 Announce Type: replace Abstract: Reinforcement Learning from Human Feedback (RLHF) relies on reward models to align large language models with human preferences. However, RLHF often

safetyarxiv-cs-lg
25 Jun 2026
Model Releases

SeFi-Image: A Text-to-Image Foundation Model with Semantic-First Diffusion

DGX agent

arXiv:2606.22568v1 Announce Type: new Abstract: Training image generation foundation models consumes substantial resources. Previous methods have attempted to leverage semantic guidance to accelerate

model-releasesarxiv-cs-cv
23 Jun 2026
Agents

Agentic Environment Engineering for Large Language Models: A Survey of Environment Modeling, Synthesis, Evaluation, and Application

DGX agent

arXiv:2606.12191v1 Announce Type: cross Abstract: Environments serve as interactive systems for large language model (LLM) based agents across diverse scenarios and play a crucial role in driving the

agentsarxiv-cs-ai
11 Jun 2026
Model Releases

BioMamba: Domain-Adaptive Biomedical Language Models

DGX agent

arXiv:2408.02600v3 Announce Type: replace Abstract: Background. Biomedical language models should improve performance on biomedical text while retaining general-language-modeling fluency. For Mamba-ba

model-releasesarxiv-cs-cl
11 Jun 2026
Model Releases

When Roleplaying, Do Models Believe What They Say?

DGX agent

arXiv:2606.11502v1 Announce Type: cross Abstract: Language models can state that 'the Earth orbits the Sun' and, when role-playing Aristotle, assert the opposite. Recent work argues that persona adopt

model-releasesarxiv-cs-ai
11 Jun 2026
Research

Large Language Models as Modal Models in Linguistics

DGX agent

arXiv:2606.10467v1 Announce Type: new Abstract: The rapid advancement of large language models (LLMs) has intensified debates about their significance for linguistic theory. These debates are commonly

researcharxiv-cs-cl
10 Jun 2026
Model Releases

Auditing Proprietary Alignment in Large Language Models: A Comparative Framework Without a Ground-Truth Standard

DGX agent

arXiv:2606.08381v1 Announce Type: cross Abstract: Large language models (LLMs) are increasingly released and deployed through opaque development and deployment pipelines, enabling model providers to i

model-releasesarxiv-cs-ai
9 Jun 2026
Research

BEACON: Behavioral Entropy Aggregation for Cross-Model Hallucination Detection in Large Language Models

DGX agent

arXiv:2606.07528v1 Announce Type: cross Abstract: Hallucination in large language models (LLMs), defined as the generation of factually incorrect or unsupported content, remains a critical barrier to

researcharxiv-cs-ai
9 Jun 2026
Applications

Cutting LLM Evaluation Costs with SySRs: A Bandit Algorithm that Provably Exploits Model Similarity

DGX agent

arXiv:2606.07726v1 Announce Type: new Abstract: Large Language Models are typically benchmarked by evaluating every model on every test query. For practitioners seeking the best model to deploy, this

applicationsarxiv-cs-lg
9 Jun 2026
Agents

GPT-Micro: A large language paradigm for accelerated, inexpensive, and thermodynamics-consistent discovery of constitutive models in manufacturing

DGX agent

arXiv:2606.08238v1 Announce Type: new Abstract: Constitutive modeling of the relationship between process-imposed material states and fundamental material properties is critical to control of material

agentsarxiv-cs-lg
9 Jun 2026
Model Releases

SLMJury: Can Small Language Models Judge as Well as Large Ones?

DGX agent

arXiv:2606.07810v1 Announce Type: cross Abstract: Large language models (LLMs) are widely used as judges for evaluating model outputs, but their high cost, latency, and opacity limit scalability. We i

model-releasesarxiv-cs-ai
9 Jun 2026
Model Releases

Characterizing Learning Dynamics under Relative Reparameterization of Singular Models

DGX agent

arXiv:2206.08598v2 Announce Type: replace Abstract: A common way to analyze learning of statistical models is to consider operations in the models parameter space, however this becomes challenging whe

model-releasesarxiv-cs-lg
8 Jun 2026
Model Releases

Data-Constrained Language Model Pretraining: Improved Regularization and Scaling Laws

DGX agent

arXiv:2606.06888v1 Announce Type: new Abstract: Classical scaling laws for language model pretraining balance model size against training dataset size under a fixed compute budget, assuming abundant d

model-releasesarxiv-cs-lg
8 Jun 2026
Local Ai

Learning Explicit Behavioral Models with Adaptive Questions and World-Model Probes

DGX agent

arXiv:2606.07127v1 Announce Type: new Abstract: Interactive agents trained only against task return can achieve high scores while failing to represent the mechanisms that make their actions succeed. T

local-aiarxiv-cs-lg
8 Jun 2026
Model Releases

Model Recycling Framework for Multi-Source Data-Free Supervised Transfer Learning

DGX agent

arXiv:2508.02039v2 Announce Type: replace Abstract: Increasing concerns for data privacy and other difficulties associated with retrieving source data for model training have created the need for sour

model-releasesarxiv-cs-lg
8 Jun 2026
Model Releases

Think Fast: Estimating No-CoT Task-Completion Time Horizons of Frontier AI Models

DGX agent

arXiv:2606.07157v1 Announce Type: new Abstract: Many efforts to ensure frontier AI models are safe rely on monitoring their chain-of-thought (CoT) reasoning. If models become able to perform sufficien

model-releasesarxiv-cs-ai
8 Jun 2026
Applications

Five labs, five minds: building a multi-model finance drama on small models

DGX agent

This article describes a collaborative hackathon project involving five research labs that developed a financial simulation drama using small language models, focusing on building complex multi-agent

applicationshugging-face
6 Jun 2026
Model Releases

LDARNet: DNA Adaptive Representation Network with Learnable Tokenization for Genomic Modeling

DGX agent

arXiv:2606.04552v1 Announce Type: new Abstract: Genomic foundation models increasingly adopt large language model architectures, yet almost universally rely on fixed tokenization schemes such as k-mer

model-releasesarxiv-cs-cl
4 Jun 2026
← Previous
1…1718192021…1238
Next →