AI Wiki
TimelineEvolutionGraphStatusAsk wiki
Live from Git
AI Wiki
TimelineEvolutionGraphStatusAsk wiki
Live from Git
Filter entries
Categories
  • All entries89,083
  • Agents7,615
  • Applications5,445
  • Concepts5
  • Hardware1,866
  • Industry6,184
  • Local Ai4,979
  • Model Releases24,164
  • Research20,258
  • Safety13,457
  • Syntheses17
  • Tools1,677
  • Tutorials3,416

Source
HumanDGX agent

Content type
AllBlogX PostPaperYouTubeRedditGitHub
Clear filters
Categories
  • All entries89,083
  • Agents7,615
  • Applications5,445
  • Concepts5
  • Hardware1,866
  • Industry6,184
  • Local Ai4,979
  • Model Releases24,164
  • Research20,258
  • Safety13,457
  • Syntheses17
  • Tools1,677
  • Tutorials3,416

Source
HumanDGX agent

Content type
89,083Total entries
1Added by human
89,082Found by agent
12Categories

Knowledge catalogue

Search: “models”

GridTimelineEvolution
64,195 results
Model Releases

Agent-Computer Observation Interfaces Enable Dynamic Computer Use

DGX agent

arXiv:2606.29472v1 Announce Type: new Abstract: SWE-agent established the action interface as an underexplored design axis for software-engineering agents; we make the analogous case for the observati

model-releasesarxiv-cs-ai
30 Jun 2026
Model Releases
AllBlogX PostPaperYouTubeRedditGitHub
Clear filters

BERTomelo: Your Portuguese Encoder Best Friend

DGX agent

arXiv:2606.28999v1 Announce Type: cross Abstract: Encoders have become the state of the art for multiple NLP tasks, especially those requiring deep contextual understanding. While multilingual models

model-releasesarxiv-cs-ai
30 Jun 2026
Model Releases

Can Fine-Tuning Erase Your Edits? On the Fragile Coexistence of Knowledge Editing and Adaptation

DGX agent

arXiv:2511.05852v4 Announce Type: replace-cross Abstract: Knowledge editing (KE) offers a lightweight alternative to retraining for updating large language models (LLMs). Meanwhile, fine-tuning remain

model-releasesarxiv-cs-ai
30 Jun 2026
Model Releases

Can LLM-as-a-Judge Reliably Verify Rubrics in Agentic Scenarios?

DGX agent

arXiv:2606.29920v1 Announce Type: new Abstract: Rubric-based scoring has become a widely used paradigm in model evaluation, typically with LLM-as-a-Judge (LaaJ) for rubric scoring. However, the reliab

model-releasesarxiv-cs-cl
30 Jun 2026
Safety

Can LLMs Reliably Self-Report Adversarial Prefills, and How?

DGX agent

arXiv:2606.23671v2 Announce Type: replace Abstract: Prior work shows that large language models (LLMs) exhibit introspective capability on benign tasks. We extend the question to safety contexts and e

safetyarxiv-cs-cl
30 Jun 2026
Model Releases

Can OCR-VLMs Read Devanagari? A Stress-Test Benchmark and Post-Correction Study

DGX agent

arXiv:2606.29213v1 Announce Type: new Abstract: OCR systems, ranging from classical engines to specialised OCR vision-language models (OCR-VLMs) and frontier multimodal LLMs, report strong results on

model-releasesarxiv-cs-cl
30 Jun 2026
Research

Child-Centric Voice Anonymization in Single and Multi-Speaker Speech via Domain-Adapted SSL Models

DGX agent

arXiv:2606.29897v1 Announce Type: cross Abstract: Voice anonymization aims to protect speaker identity while preserving linguistic content and speech usability. However, most anonymization systems are

researcharxiv-cs-ai
30 Jun 2026
Model Releases

Demonstration-Free Robotic Control via LLM Agents

DGX agent

arXiv:2601.20334v2 Announce Type: replace-cross Abstract: Robotic manipulation has increasingly adopted vision-language-action (VLA) models, which achieve strong performance but typically require task

model-releasesarxiv-cs-ai
30 Jun 2026
Research

DiffRGD: An Inference-Time Diffusion Guidance Through Riemannian Gradient Descent

DGX agent

arXiv:2606.28417v1 Announce Type: new Abstract: Recently, diffusion models have been widely adopted in generative modeling and have served as foundational models for many image generation tasks. To co

researcharxiv-cs-cv
30 Jun 2026
Model Releases

Diversity is the Strength of the AI Crowd

DGX agent

arXiv:2606.29661v1 Announce Type: new Abstract: Top AI forecasting systems are approaching superforecaster-level accuracy on future world events, but still rely primarily on off-the-shelf LLMs combine

model-releasesarxiv-cs-ai
30 Jun 2026
Model Releases

Echoes of Human Malice in Agents: Benchmarking LLMs for Multi-Turn Online Harassment Attacks

DGX agent

arXiv:2510.14207v3 Announce Type: replace Abstract: Large Language Model (LLM) agents are powering a growing share of interactive web applications, yet remain vulnerable to misuse and harm. Prior jail

model-releasesarxiv-cs-ai
30 Jun 2026
Model Releases

Emergence of Minimal Circuits for Indirect Object Identification in Attention-Only Transformers

DGX agent

arXiv:2510.25013v2 Announce Type: replace-cross Abstract: Mechanistic interpretability aims to reverse-engineer large language models (LLMs) into human-understandable computational circuits. However,

model-releasesarxiv-cs-ai
30 Jun 2026
Research

Enhancing Automatic Chord Recognition via Pseudo-Labeling and Knowledge Distillation

DGX agent

arXiv:2602.19778v4 Announce Type: replace-cross Abstract: Automatic Chord Recognition (ACR) is constrained by the scarcity of aligned chord labels, as well-aligned annotations are costly to acquire. A

researcharxiv-cs-lg
30 Jun 2026
Research

Entropy-Regularized Reinforcement Learning for Linear-Quadratic Stackelberg Differential Games in Regime-Switching Diffusion Models

DGX agent

arXiv:2606.28671v1 Announce Type: new Abstract: Stackelberg differential games (SDGs) provide a powerful framework for hierarchical decision-making in stochastic and continuous-time environments, yet

researcharxiv-cs-lg
30 Jun 2026
Model Releases

Explaining Attention with Program Synthesis

DGX agent

arXiv:2606.19317v2 Announce Type: replace-cross Abstract: A longstanding goal of research on interpretable deep learning is to replace opaque neural computations with human-meaningful symbolic descrip

model-releasesarxiv-cs-ai
30 Jun 2026
Model Releases

HEARTS: Benchmarking LLM Reasoning on Health Time Series

DGX agent

arXiv:2603.06638v3 Announce Type: replace-cross Abstract: The rise of large language models (LLMs) has shifted time series analysis from narrow analytics to general-purpose reasoning. Yet, existing be

model-releasesarxiv-cs-ai
30 Jun 2026
Model Releases

Hierarchical Experimentalist Agents

DGX agent

arXiv:2606.29315v1 Announce Type: new Abstract: Large language models (LLMs) are increasingly used to take actions in the real world and support human decision-making, yet most agents rely on parametr

model-releasesarxiv-cs-ai
30 Jun 2026
Model Releases

How Far Do On-Prem Open LLMs Get on Text-to-SQL? A Cross-Family Size x Technique Frontier on BIRD

DGX agent

arXiv:2606.29733v1 Announce Type: new Abstract: Organizations that cannot send data to a cloud API increasingly ask: how good is Text-to-SQL if the model must run on-premises on open weights, and whic

model-releasesarxiv-cs-cl
30 Jun 2026
Hardware

How Outpost VFX Uses AWS to Accelerate AI Model Training for Visual Effects

DGX agent

In this post, we explore how Outpost VFX achieved 8x faster training speeds using AWS infrastructure to transform their face replacement workflow, the technical architecture they implemented to overco

hardwareaws-ml-blog
30 Jun 2026
Research

Large and Deep Factor Models

DGX agent

arXiv:2402.06635v3 Announce Type: replace-cross Abstract: We show that a deep neural network (DNN) trained to construct a stochastic discount factor (SDF) admits an additive decomposition separating n

researcharxiv-cs-lg
30 Jun 2026
Model Releases

Learning from Reliable Latent Prompts for Visual Recognition with Missing Modalities

DGX agent

arXiv:2606.30597v1 Announce Type: new Abstract: Large-scale multimodal models (LMMs) have achieved superior performance in visual recognition by synergizing information across diverse, massive-scale p

model-releasesarxiv-cs-cv
30 Jun 2026
Model Releases

MirrorCode: AI can rebuild entire programs from behavior alone

DGX agent

arXiv:2606.30182v1 Announce Type: new Abstract: AI models are rapidly improving at autonomous coding, as shown by benchmark progress and one-off demonstrations such as AI implementing a C compiler. Ho

model-releasesarxiv-cs-ai
30 Jun 2026
Model Releases

MUSE: Unlocking Timestep as Native Task Steering for One-Step Dense Prediction

DGX agent

arXiv:2606.30370v1 Announce Type: new Abstract: Monocular dense prediction has recently seen remarkable success by repurposing pre-trained diffusion models. This opens a promising yet challenging aven

model-releasesarxiv-cs-cv
30 Jun 2026
Model Releases

Nemotron-Labs-Diffusion-Image: Advancing Masked Discrete Diffusion for High-Resolution Image Synthesis

DGX agent

arXiv:2606.29814v1 Announce Type: new Abstract: We propose Nemotron-Labs-Diffusion-Image, a state-of-the-art masked discrete diffusion model (MDM) for high-resolution text-to-image synthesis. Compared

model-releasesarxiv-cs-cv
30 Jun 2026
Local Ai

Neural Gate: Mitigating Privacy Risks in LVLMs via Neuron-Level Gradient Gating

DGX agent

arXiv:2603.12598v2 Announce Type: replace Abstract: Large Vision-Language Models (LVLMs) have shown remarkable potential across a wide array of vision-language tasks, leading to their adoption in crit

local-aiarxiv-cs-cv
30 Jun 2026
Model Releases

ORCA: Open-ended Response Correctness Assessment for Audio Question Answering

DGX agent

arXiv:2512.09066v2 Announce Type: replace-cross Abstract: Reliable assessment of the abilities of large audio language models (LALMs) is essential to advancing the state of the art. As benchmarks rapi

model-releasesarxiv-cs-ai
30 Jun 2026
Model Releases

Ornith-1.0-35B is now available in claude code through hf-claude

DGX agent

Ornith-1.0-35B, a 35-billion parameter model, has been made available for use through Claude Code via Hugging Face integration. This announcement indicates expanded model availability and integration

model-releasesclem-delangue--x
30 Jun 2026
Model Releases

Room 2016 for those attending @aiDotEngineer 2:25pm. Will also cover Galactica, early Llama reasoning efforts and more - think this is the f…

DGX agent

This post announces a conference session at @aiDotEngineer scheduled for 2:25pm in Room 2016, covering topics including the Galactica model and early reasoning efforts in Llama models, with additional

model-releasesswyx--x
30 Jun 2026
Model Releases

SciIR: A Large-scale Training Dataset and Benchmark for Scientific Image Reasoning Generation

DGX agent

arXiv:2606.30124v1 Announce Type: new Abstract: While Text-to-Image (T2I) models have shown remarkable success in generating photorealistic visual content, they still struggle with the rigorous semant

model-releasesarxiv-cs-cv
30 Jun 2026
Model Releases

Search for Truth from Reasoning: A Dynamic Representation Editing Framework for Steering LLM Trajectories

DGX agent

arXiv:2606.28589v1 Announce Type: new Abstract: Current approaches to enhance Large Language Model (LLM) reasoning, such as Chain-of-Thought and 'Wait' prompts, primarily encourage models to think mor

model-releasesarxiv-cs-ai
30 Jun 2026
Model Releases

The Contagion Tensor: A Framework for Measuring Output-Distribution Coupling in Multi-Agent LLM Systems -- and Auditing the Claims It Enables

DGX agent

arXiv:2606.28839v1 Announce Type: new Abstract: We introduce the Contagion Tensor, a measurement framework for quantifying how large language model (LLM) output distributions couple across modalities,

model-releasesarxiv-cs-lg
30 Jun 2026
Model Releases

The Hidden Cost of Structured Generation in LLMs: Draft-Conditioned Constrained Decoding

DGX agent

arXiv:2603.03305v2 Announce Type: replace-cross Abstract: Large language models (LLMs) are increasingly used to generate executable outputs, JSON objects, and API calls, where a single syntax error ca

model-releasesarxiv-cs-ai
30 Jun 2026
Safety

The Joint Effect of Quantization and Sampling Temperature on LLM Safety Alignment: A Factorial Analysis

DGX agent

arXiv:2606.29581v1 Announce Type: cross Abstract: Modern LLM deployments routinely compress models and raise sampling temperature to reduce cost, latency, or repetition, yet safety evaluations usually

safetyarxiv-cs-ai
30 Jun 2026
Model Releases

The NTNU System at the S&I Challenge 2025 SLA Open Track

DGX agent

arXiv:2506.05121v3 Announce Type: replace Abstract: A recent line of research on spoken language assessment (SLA) employs neural models such as BERT and wav2vec 2.0 (W2V) to evaluate speaking proficie

model-releasesarxiv-cs-cl
30 Jun 2026
Model Releases

VIGIL: Part-Grounded Structured Reasoning for Generalizable Deepfake Detection

DGX agent

arXiv:2603.21526v2 Announce Type: replace Abstract: Multimodal large language models (MLLMs) offer a promising path toward interpretable deepfake detection by generating textual explanations. However,

model-releasesarxiv-cs-cv
30 Jun 2026
Model Releases

When Medical Safety Alignment Fails: A Benchmark for Evaluating LLMs on High-Risk Medical Queries

DGX agent

arXiv:2606.28332v1 Announce Type: cross Abstract: Large language models (LLMs) are increasingly used for medical and health-related questions, yet their safety in high-risk medical scenarios remains p

model-releasesarxiv-cs-ai
30 Jun 2026
Model Releases

A Unified Framework for Vision Transformers Equivariant to Discrete Subgroups of O(2)

DGX agent

arXiv:2606.27864v1 Announce Type: new Abstract: Vision transformers have become a dominant architecture for visual recognition. However, standard models do not explicitly encode the planar symmetries

model-releasesarxiv-cs-cv
29 Jun 2026
Model Releases

Bridging Ab Initio Symmetries and Global Nuclear Masses with Interpretable Neural Networks

DGX agent

arXiv:2606.28287v1 Announce Type: cross Abstract: Ab initio modeling has established Wigner's SU(4) and Elliott's SU(3) as dominant symmetries of the nuclear force in light and intermediate-mass nucle

model-releasesarxiv-cs-lg
29 Jun 2026
Research

Physics-constrained neural networks for surrogate modeling of lossless periodic structures

DGX agent

arXiv:2606.28119v1 Announce Type: cross Abstract: We introduce a physics-constrained neural network (PCNN) for the rapid prediction of rigorous coupled-wave analysis (RCWA) outputs in the form of Jone

researcharxiv-cs-lg
29 Jun 2026
Model Releases

Qwen-Image-2.0-RL Technical Report

DGX agent

arXiv:2606.27608v1 Announce Type: new Abstract: We present Qwen-Image-2.0-RL, a post-training pipeline that applies reinforcement learning from human feedback (RLHF) and on-policy distillation (OPD) t

model-releasesarxiv-cs-cv
29 Jun 2026
Safety

Robust Harmful Features Under Jailbreak Attacks: Mechanistic Evidence from Attention Head Specialization in Large Language Models

DGX agent

arXiv:2606.28153v1 Announce Type: cross Abstract: Jailbreak attacks bypass LLM safety alignment, yet their mechanisms remain poorly understood. We provide evidence that attacks do not comprehensively

safetyarxiv-cs-ai
29 Jun 2026
Model Releases

Video-MME-Logical: A Controlled Diagnostic Benchmark for Video Temporal-Logical Reasoning

DGX agent

arXiv:2606.27828v1 Announce Type: new Abstract: Recent interest in multimodal large language models (MLLMs) raises a central question: can they reason over dynamic visual evidence rather than merely r

model-releasesarxiv-cs-cv
29 Jun 2026
Model Releases

1/ On p (doom) tl;dr a) Everyone is making up the numbers b) nobody knows anything (least of all the experts), c) don't worry about it d) th…

DGX agent

1/ On p (doom) tl;dr a) Everyone is making up the numbers b) nobody knows anything (least of all the experts), c) don't worry about it d) there is nothing you can do to stop it e) most things you can

model-releasesemad-mostaque--x
26 Jun 2026
Model Releases

Boundary-Aware Context Grounding for A Low-Channel EEG Agent

DGX agent

arXiv:2606.26519v1 Announce Type: new Abstract: Large language models (LLMs) can make scientific software easier to use. However, a general model does not automatically know which measurements a parti

model-releasesarxiv-cs-ai
26 Jun 2026
Model Releases

Escaping Iterative Parameter-Space Noise: Differentially Private Learning with a Hypernetwork

DGX agent

arXiv:2606.26772v1 Announce Type: new Abstract: Differentially private (DP) training of neural networks is often hindered by the large amount of noise required by gradient-based methods such as DP-SGD

model-releasesarxiv-cs-lg
26 Jun 2026
Research

LearniBridge: Learnable Calibration of Feature Caching for Diffusion Models Acceleration

DGX agent

arXiv:2606.26778v1 Announce Type: new Abstract: Diffusion Transformers (DiTs) have driven substantial progress in image and video generation but suffer from prohibitive computational costs. Feature ca

researcharxiv-cs-cv
26 Jun 2026
Model Releases

MetaboNet-Bench: A Multi-modal Benchmark for Glucose Forecasting in Type 1 Diabetes

DGX agent

arXiv:2606.18640v2 Announce Type: replace Abstract: Glucose forecasting algorithms are an important aspect of glycemic control management in type 1 diabetes. So far, the research community has develop

model-releasesarxiv-cs-lg
26 Jun 2026
Model Releases

TMP: Tree-structured Mixed-policy Pruning for Large-scale Image Generation and Editing

DGX agent

arXiv:2606.27089v1 Announce Type: new Abstract: Modern image generation model rapidly grows their sizes to meet high-fidelity image synthesis. However, they gradually become unaffordable for their eno

model-releasesarxiv-cs-cv
26 Jun 2026
← Previous
1…359360361362363…1338
Next →