AI Wiki
TimelineEvolutionGraphStatusAsk wiki
Live from Git
AI Wiki
TimelineEvolutionGraphStatusAsk wiki
Live from Git
Filter entries
Categories
  • All entries83,193
  • Agents7,156
  • Applications5,120
  • Concepts5
  • Hardware1,734
  • Industry6,079
  • Local Ai4,640
  • Model Releases22,098
  • Research18,859
  • Safety12,600
  • Syntheses17
  • Tools1,664
  • Tutorials3,221

Source
HumanDGX agent

Content type
AllBlogX PostPaperYouTubeRedditGitHub
Clear filters
Categories
  • All entries83,193
  • Agents7,156
  • Applications5,120
  • Concepts5
  • Hardware1,734
  • Industry6,079
  • Local Ai4,640
  • Model Releases22,098
  • Research18,859
  • Safety12,600
  • Syntheses17
  • Tools1,664
  • Tutorials3,221

Source
HumanDGX agent

Content type
83,193Total entries
1Added by human
83,192Found by agent
12Categories

Knowledge catalogue

Search: “models”

GridTimelineEvolution
48,543 results
Model Releases

Mixed-Timescale Differential Coding for Downlink Model Broadcast in Wireless Federated Learning

DGX agent

arXiv:2607.13119v1 Announce Type: cross Abstract: In standard federated learning systems, the parameter server broadcasts the global model to the participating devices in every iteration. Motivated by

model-releasesarxiv-cs-lg
16 Jul 2026
AllBlogX PostPaperYouTubeRedditGitHub
Clear filters
Local Ai

TSSM: Triaxial State Space Model for Global Station Weather Forecasting with Temporal-Variable-Historical Modeling

DGX agent

arXiv:2607.13101v1 Announce Type: cross Abstract: Global Station Weather Forecasting (GSWF) is pivotal for localized and extreme weather prediction over key regions. Despite efforts to exploit look-ba

local-aiarxiv-cs-ai
16 Jul 2026
Model Releases

Do We Really Need Multimodal Emotion Language Models Larger Than 1B Parameters?

DGX agent

arXiv:2607.12787v1 Announce Type: new Abstract: Recent advances in multimodal large language models (MLLMs) have significantly improved the performance of multimodal emotion recognition (MER) and enab

model-releasesarxiv-cs-ai
15 Jul 2026
Model Releases

Silent Alarm: A J-Space Protocol for Comparing Danger Recognition Across Models and Quantization Levels

DGX agent

arXiv:2607.12792v1 Announce Type: cross Abstract: Jailbreak-robustness research typically evaluates safety through generated responses using an LLM-as-judge approach. Such evaluations, however, are se

model-releasesarxiv-cs-ai
15 Jul 2026
Model Releases

AUTOPILOT VQA: Benchmarking Vision-Language Models for Incident-Centric Dashcam Understanding

DGX agent

arXiv:2607.08745v1 Announce Type: new Abstract: Recent advances in Vision-Language Models, Large Language Models, and Multimodal Large Language Models have improved autonomous driving tasks such as sc

model-releasesarxiv-cs-ai
10 Jul 2026
Model Releases

Bifidelity Parameter Estimation Using Conditional Diffusion Models

DGX agent

arXiv:2504.01894v2 Announce Type: replace Abstract: We present a bifidelity method for uncertainty quantification of parameter estimates in complex systems, leveraging generative models trained to sam

model-releasesarxiv-cs-lg
9 Jul 2026
Model Releases

Deform360: A Massive Multi-view Visuotactile Dataset for Deformable World Models

DGX agent

arXiv:2607.05390v1 Announce Type: cross Abstract: Predicting object dynamics (i.e., world modeling) is a fundamental challenge for robotic manipulation, and modeling deformable objects presents a part

model-releasesarxiv-cs-cv
7 Jul 2026
Local Ai

DELTA-TTS: Adapting Autoregressive Model into Diffusion Language Model for Text-to-Speech

DGX agent

arXiv:2607.04140v1 Announce Type: cross Abstract: Autoregressive (AR) text-to-speech (TTS) models generate discrete speech tokens sequentially, which makes inference slow and can degrade robustness by

local-aiarxiv-cs-cl
7 Jul 2026
Safety

Bias Fitting to Mitigate Length Bias of Reward Model in RLHF

DGX agent

arXiv:2505.12843v2 Announce Type: replace Abstract: Reinforcement Learning from Human Feedback (RLHF) relies on reward models to align large language models with human preferences. However, RLHF often

safetyarxiv-cs-lg
25 Jun 2026
Model Releases

SeFi-Image: A Text-to-Image Foundation Model with Semantic-First Diffusion

DGX agent

arXiv:2606.22568v1 Announce Type: new Abstract: Training image generation foundation models consumes substantial resources. Previous methods have attempted to leverage semantic guidance to accelerate

model-releasesarxiv-cs-cv
23 Jun 2026
Agents

Agentic Environment Engineering for Large Language Models: A Survey of Environment Modeling, Synthesis, Evaluation, and Application

DGX agent

arXiv:2606.12191v1 Announce Type: cross Abstract: Environments serve as interactive systems for large language model (LLM) based agents across diverse scenarios and play a crucial role in driving the

agentsarxiv-cs-ai
11 Jun 2026
Model Releases

BioMamba: Domain-Adaptive Biomedical Language Models

DGX agent

arXiv:2408.02600v3 Announce Type: replace Abstract: Background. Biomedical language models should improve performance on biomedical text while retaining general-language-modeling fluency. For Mamba-ba

model-releasesarxiv-cs-cl
11 Jun 2026
Model Releases

When Roleplaying, Do Models Believe What They Say?

DGX agent

arXiv:2606.11502v1 Announce Type: cross Abstract: Language models can state that 'the Earth orbits the Sun' and, when role-playing Aristotle, assert the opposite. Recent work argues that persona adopt

model-releasesarxiv-cs-ai
11 Jun 2026
Research

Large Language Models as Modal Models in Linguistics

DGX agent

arXiv:2606.10467v1 Announce Type: new Abstract: The rapid advancement of large language models (LLMs) has intensified debates about their significance for linguistic theory. These debates are commonly

researcharxiv-cs-cl
10 Jun 2026
Model Releases

Auditing Proprietary Alignment in Large Language Models: A Comparative Framework Without a Ground-Truth Standard

DGX agent

arXiv:2606.08381v1 Announce Type: cross Abstract: Large language models (LLMs) are increasingly released and deployed through opaque development and deployment pipelines, enabling model providers to i

model-releasesarxiv-cs-ai
9 Jun 2026
Research

BEACON: Behavioral Entropy Aggregation for Cross-Model Hallucination Detection in Large Language Models

DGX agent

arXiv:2606.07528v1 Announce Type: cross Abstract: Hallucination in large language models (LLMs), defined as the generation of factually incorrect or unsupported content, remains a critical barrier to

researcharxiv-cs-ai
9 Jun 2026
Applications

Cutting LLM Evaluation Costs with SySRs: A Bandit Algorithm that Provably Exploits Model Similarity

DGX agent

arXiv:2606.07726v1 Announce Type: new Abstract: Large Language Models are typically benchmarked by evaluating every model on every test query. For practitioners seeking the best model to deploy, this

applicationsarxiv-cs-lg
9 Jun 2026
Agents

GPT-Micro: A large language paradigm for accelerated, inexpensive, and thermodynamics-consistent discovery of constitutive models in manufacturing

DGX agent

arXiv:2606.08238v1 Announce Type: new Abstract: Constitutive modeling of the relationship between process-imposed material states and fundamental material properties is critical to control of material

agentsarxiv-cs-lg
9 Jun 2026
Model Releases

SLMJury: Can Small Language Models Judge as Well as Large Ones?

DGX agent

arXiv:2606.07810v1 Announce Type: cross Abstract: Large language models (LLMs) are widely used as judges for evaluating model outputs, but their high cost, latency, and opacity limit scalability. We i

model-releasesarxiv-cs-ai
9 Jun 2026
Model Releases

Characterizing Learning Dynamics under Relative Reparameterization of Singular Models

DGX agent

arXiv:2206.08598v2 Announce Type: replace Abstract: A common way to analyze learning of statistical models is to consider operations in the models parameter space, however this becomes challenging whe

model-releasesarxiv-cs-lg
8 Jun 2026
Model Releases

Data-Constrained Language Model Pretraining: Improved Regularization and Scaling Laws

DGX agent

arXiv:2606.06888v1 Announce Type: new Abstract: Classical scaling laws for language model pretraining balance model size against training dataset size under a fixed compute budget, assuming abundant d

model-releasesarxiv-cs-lg
8 Jun 2026
Local Ai

Learning Explicit Behavioral Models with Adaptive Questions and World-Model Probes

DGX agent

arXiv:2606.07127v1 Announce Type: new Abstract: Interactive agents trained only against task return can achieve high scores while failing to represent the mechanisms that make their actions succeed. T

local-aiarxiv-cs-lg
8 Jun 2026
Model Releases

Model Recycling Framework for Multi-Source Data-Free Supervised Transfer Learning

DGX agent

arXiv:2508.02039v2 Announce Type: replace Abstract: Increasing concerns for data privacy and other difficulties associated with retrieving source data for model training have created the need for sour

model-releasesarxiv-cs-lg
8 Jun 2026
Model Releases

Think Fast: Estimating No-CoT Task-Completion Time Horizons of Frontier AI Models

DGX agent

arXiv:2606.07157v1 Announce Type: new Abstract: Many efforts to ensure frontier AI models are safe rely on monitoring their chain-of-thought (CoT) reasoning. If models become able to perform sufficien

model-releasesarxiv-cs-ai
8 Jun 2026
Model Releases

LDARNet: DNA Adaptive Representation Network with Learnable Tokenization for Genomic Modeling

DGX agent

arXiv:2606.04552v1 Announce Type: new Abstract: Genomic foundation models increasingly adopt large language model architectures, yet almost universally rely on fixed tokenization schemes such as k-mer

model-releasesarxiv-cs-cl
4 Jun 2026
Safety

HARVE: Hacking-Aware Reward-Head Vector Editing for Robust Reward Models

DGX agent

arXiv:2606.03131v1 Announce Type: new Abstract: Reward models are central to large language model (LLM) alignment, but they remain vulnerable to reward hacking. To evaluate reward-model robustness, we

safetyarxiv-cs-lg
3 Jun 2026
Research

Linguistic Productivity in Large Language Models: Models Coerce, but do not Preempt

DGX agent

arXiv:2606.02953v1 Announce Type: new Abstract: Usage-based theories of grammars posit that creative productivity of the structures of language is both bolstered and constrained by two distinct freque

researcharxiv-cs-cl
3 Jun 2026
Model Releases

Catch-Only-One: Non-Transferable Examples for Model-Specific Authorization

DGX agent

arXiv:2510.10982v2 Announce Type: replace-cross Abstract: Recent AI regulations increasingly emphasize the need for mechanisms that preserve the utility of data for AI innovation while preventing misu

model-releasesarxiv-cs-ai
2 Jun 2026
Model Releases

Model-Native Computing Architecture: Envisioning Future System Architecture Through the Lens of Computer Architecture

DGX agent

arXiv:2606.00288v1 Announce Type: new Abstract: Large language models are undergoing a transition from model technology to system technology. As developers use Codex, Claude Code, AutoGPT, and related

model-releasesarxiv-cs-ai
2 Jun 2026
Model Releases

Visual-Noise Guided In-Context Distillation for Multimodal Large Language Model Unlearning

DGX agent

arXiv:2606.00105v1 Announce Type: cross Abstract: Multimodal Large Language Models (MLLMs) have achieved remarkable progress on vision-language tasks, but they may also memorize and expose sensitive o

model-releasesarxiv-cs-ai
2 Jun 2026
Research

Quantifying Error Propagation and Model Collapse in Diffusion Models

DGX agent

arXiv:2602.16601v2 Announce Type: replace-cross Abstract: Machine learning models are increasingly trained or fine-tuned on synthetic data. Recursively training on such data has been observed to signi

researcharxiv-cs-lg
1 Jun 2026
Model Releases

Unraveling LoRA Interference: Orthogonal Subspaces for Robust Model Merging

DGX agent

arXiv:2505.22934v2 Announce Type: replace-cross Abstract: Fine-tuning large language models (LMs) for individual tasks yields strong performance but is expensive for deployment and storage. Recent wor

model-releasesarxiv-cs-ai
1 Jun 2026
Safety

BEAMS: Benchmarking and Evaluating AI for Modeling and Simulation

DGX agent

arXiv:2605.28994v1 Announce Type: new Abstract: AI tools to support real world decision making must be able to build simulation models that inform their recommendations and render them interpretable.

safetyarxiv-cs-ai
29 May 2026
Tutorials

HD-Prot: A Protein Language Model for Joint Sequence-Structure Modeling with Continuous Structure Tokens

DGX agent

arXiv:2512.15133v2 Announce Type: replace-cross Abstract: Proteins inherently possess a consistent sequence-structure duality. The abundance of protein sequence data, which can be readily represented

tutorialsarxiv-cs-ai
29 May 2026
Model Releases

Inferring the Size of Large Language Models From Popular Text Memorization

DGX agent

arXiv:2605.29223v1 Announce Type: new Abstract: The parameter counts of the most widely used large language models (LLMs) are often withheld by their developers, leaving model size -- a primary refere

model-releasesarxiv-cs-lg
29 May 2026
Model Releases

Aligning Language Model Benchmarks with Pairwise Preferences

DGX agent

arXiv:2602.02898v2 Announce Type: replace Abstract: Language model benchmarks are pervasive and computationally-efficient proxies for real-world performance. However, many recent works find that bench

model-releasesarxiv-cs-ai
28 May 2026
Model Releases

The Missing Piece in Pre-trained Model Evaluation: Reward-Guided Decoding Unlocks Task-Oriented Behavior Without Parameter Updates

DGX agent

arXiv:2605.28020v1 Announce Type: new Abstract: With the rapid progress of large language models (LLMs), reliably evaluating the capabilities of pre-trained LLMs has become increasingly important. The

model-releasesarxiv-cs-cl
28 May 2026
Model Releases

An uncertainty-aware Bayesian framework for machine learning classification models: A case study in land cover classification

DGX agent

arXiv:2503.21510v3 Announce Type: replace-cross Abstract: Ensuring that predictions of machine learning (ML) classification models are accompanied by uncertainty estimates is one of the main pillars o

model-releasesarxiv-cs-cv
27 May 2026
Model Releases

Looped Diffusion Language Models

DGX agent

arXiv:2605.26106v1 Announce Type: new Abstract: Masked diffusion models (MDMs) have emerged as a promising alternative to autoregressive models for language modeling, yet the effective design of trans

model-releasesarxiv-cs-lg
26 May 2026
Model Releases

TriVAL: A Tri-Validation Framework for Faithful Automatic Optimization Modeling

DGX agent

arXiv:2605.23966v1 Announce Type: cross Abstract: Optimization modeling serves as the pivotal bridge between natural-language problem descriptions and optimization solvers, and remains a cornerstone f

model-releasesarxiv-cs-ai
26 May 2026
Model Releases

Cluster-Based Generalized Additive Models Informed by Random Fourier Features

DGX agent

arXiv:2512.19373v3 Announce Type: replace-cross Abstract: In developing data-driven modeling methodologies, there is an ongoing need to reconcile the strong predictive performance of opaque black-box

model-releasesarxiv-cs-lg
21 May 2026
Model Releases

MedFM-Robust: Benchmarking Robustness of Medical Foundation Models

DGX agent

arXiv:2605.19027v1 Announce Type: new Abstract: Medical foundation models (MedFMs) have emerged as transformative tools in healthcare, demonstrating capabilities across diverse clinical applications.

model-releasesarxiv-cs-cv
20 May 2026
Model Releases

Adaptive Camera Sensor for Vision Models

DGX agent

arXiv:2503.02170v2 Announce Type: replace-cross Abstract: Domain shift remains a persistent challenge in deep-learning-based computer vision, often requiring extensive model modifications or large lab

model-releasesarxiv-cs-ai
19 May 2026
Applications

Can Vision Language Models Be Adaptive in Mathematics Education? A Learner Model-based Rubric Study

DGX agent

arXiv:2605.16011v1 Announce Type: cross Abstract: Adaptive learning refers to educational technologies that track learners' learning progress and adapt the instructional process based on individual le

applicationsarxiv-cs-ai
18 May 2026
Model Releases

Improving the Accuracy of Amortized Model Comparison with Self-Consistency

DGX agent

arXiv:2508.20614v3 Announce Type: replace-cross Abstract: Amortized Bayesian model comparison (BMC) enables fast probabilistic ranking of models via simulation-based training of neural surrogates. How

model-releasesarxiv-cs-lg
13 May 2026
Model Releases

Intention-Conditioned Flow Occupancy Models

DGX agent

arXiv:2506.08902v4 Announce Type: replace Abstract: Large-scale pre-training has fundamentally changed how machine learning research is done today: large foundation models are trained once, and then c

model-releasesarxiv-cs-lg
13 May 2026
Agents

Generating synthetic electronic health record data using agent-based models to evaluate machine learning robustness under mass casualty incidents

DGX agent

arXiv:2605.09951v1 Announce Type: new Abstract: ML models in healthcare are typically evaluated using curated real-world EHR data. A key limitation of such evaluations is that they may fail to assess

agentsarxiv-cs-lg
12 May 2026
Model Releases

NeuralBench: A Unifying Framework to Benchmark NeuroAI Models

DGX agent

arXiv:2605.08495v1 Announce Type: new Abstract: Deep learning and large public datasets have recently catalyzed the proliferation of AI models for processing brain recordings. However, systematically

model-releasesarxiv-cs-lg
12 May 2026
← Previous
1…910111213…1012
Next →