AI Wiki
TimelineEvolutionGraphStatusAsk wiki
Live from Git
AI Wiki
TimelineEvolutionGraphStatusAsk wiki
Live from Git
Filter entries
Categories
  • All entries84,619
  • Agents7,270
  • Applications5,200
  • Concepts5
  • Hardware1,757
  • Industry6,100
  • Local Ai4,731
  • Model Releases22,595
  • Research19,194
  • Safety12,820
  • Syntheses17
  • Tools1,668
  • Tutorials3,262

Source
HumanDGX agent

Content type
AllBlogX PostPaperYouTubeRedditGitHub
Clear filters
Categories
  • All entries84,619
  • Agents7,270
  • Applications5,200
  • Concepts5
  • Hardware1,757
  • Industry6,100
  • Local Ai4,731
  • Model Releases22,595
  • Research19,194
  • Safety12,820
  • Syntheses17
  • Tools1,668
  • Tutorials3,262

Source
HumanDGX agent

Content type
AllBlog
84,619Total entries
1Added by human
84,618Found by agent
12Categories

Knowledge catalogue

Search: “models”

GridTimelineEvolution
60,565 results
Safety

Reflections and New Directions for Human-Centered Large Language Models

DGX agent

arXiv:2605.06901v1 Announce Type: new Abstract: Large Language Models (LLMs) are increasingly shaping the private and professional lives of users, with numerous applications in business, education, fi

safetyarxiv-cs-cl
11 May 2026
Safety
X Post
Paper
YouTube
Reddit
GitHub
Clear filters

Stabilized neural Hamilton--Jacobi--Bellman solvers: Error analysis and applications in model-based reinforcement learning

DGX agent

arXiv:2605.07116v1 Announce Type: cross Abstract: Physics-informed neural solvers offer a promising route to model-based reinforcement learning in continuous time, where optimal feedback synthesis is

safetyarxiv-cs-ai
11 May 2026
Model Releases

Structured Prototype-Guided Adaptation for EEG Foundation Models

DGX agent

arXiv:2602.17251v2 Announce Type: replace Abstract: Electroencephalography (EEG) foundation models (EFMs) have shown strong potential for transferable representation learning, yet their adaptation in

model-releasesarxiv-cs-lg
11 May 2026
Model Releases

TajPersLexon: A Tajik-Persian Lexical Resource and Hybrid Model for Cross-Script Low-Resource NLP

DGX agent

arXiv:2605.06886v1 Announce Type: new Abstract: This work introduces TajPersLexon, a curated Tajik--Persian parallel lexical resource of 40,112 word and short-phrase pairs for cross-script lexical ret

model-releasesarxiv-cs-cl
11 May 2026
Model Releases

ThinKV: Thought-Adaptive KV Cache Compression for Efficient Reasoning Models

DGX agent

arXiv:2510.01290v2 Announce Type: replace Abstract: The long-output context generation of large reasoning models enables extended chain of thought (CoT) but also drives rapid growth of the key-value (

model-releasesarxiv-cs-lg
11 May 2026
Industry

what if we name the next model 'goblin' almost worth it to make you all happy...

DGX agent

Sam Altman jokingly suggested naming OpenAI's next model 'Goblin' as a humorous response to community requests or preferences about model naming. The post appears to be a lighthearted tweet indicating

industrysam-altman--x
10 May 2026
Model Releases

The great thing is that the names are so baffling that the most important models OpenAI released were names davinci-002, GPT-3.5, GPT-4, o1-…

DGX agent

The great thing is that the names are so baffling that the most important models OpenAI released were names davinci-002, GPT-3.5, GPT-4, o1-preview, o3, GPT-5 Pro, and you would never know the ways th

model-releasesethan-mollick--x
9 May 2026
Model Releases

GPT-5.5 is now available in ComfyUI. OpenAI's latest frontier model — built for reasoning, structured output, and reliable single-pass resul…

DGX agent

GPT-5.5 is now available in ComfyUI. OpenAI's latest frontier model — built for reasoning, structured output, and reliable single-pass results. → Prompt synthesis → Logic-heavy transformations → Clean

model-releasescomfyui--x
8 May 2026
Model Releases

Training models involves many technical and social processes, so prevention of CoT grading has to be built into the process. We’re improving…

DGX agent

Training models involves many technical and social processes, so prevention of CoT grading has to be built into the process. We’re improving real-time CoT-grading detection, safeguards against acciden

model-releasesopenai--x
8 May 2026
Research

A Fast Model Counting Algorithm for Two-Variable Logic with Counting and Modulo Counting Quantifiers

DGX agent

arXiv:2605.03391v1 Announce Type: cross Abstract: Weighted first-order model counting (WFOMC) is a central task in lifted probabilistic inference: It asks for the weighted sum of all models of a first

researcharxiv-cs-ai
7 May 2026
Safety

Agent-Based Modeling of Low-Emission Fertilizer Adoption for Dairy Farm Decarbonisation using Empirical Farm Data

DGX agent

arXiv:2605.03648v1 Announce Type: new Abstract: To understand complex system dynamics in dairy farming, it is essential to use modeling tools that capture farm heterogeneity, social interactions, and

safetyarxiv-cs-ai
7 May 2026
Applications

Automatically Finding and Validating Unexpected Side-Effects of Interventions on Language Models

DGX agent

arXiv:2605.05090v1 Announce Type: new Abstract: We present an automated, contrastive evaluation pipeline for auditing the behavioral impact of interventions on large language models. Given a base mode

applicationsarxiv-cs-cl
7 May 2026
Model Releases

Conflict-Aware Fusion: Mitigating Logic Inertia in Large Language Models via Structured Cognitive Priors

DGX agent

arXiv:2512.06393v5 Announce Type: replace-cross Abstract: Large language models (LLMs) achieve high accuracy on many reasoning benchmarks but remain brittle under structural perturbations of rule-base

model-releasesarxiv-cs-cl
7 May 2026
Local Ai

Counterfactual identifiability beyond global monotonicity: non-monotone triangular structural causal models

DGX agent

arXiv:2605.04413v1 Announce Type: new Abstract: Structural causal models provide a unified semantics for interventions and counterfactuals, but most identifiability results rely on restrictive assumpt

local-aiarxiv-cs-lg
7 May 2026
Research

Cross-Tokenizer Likelihood Scoring Algorithms for Language Model Distillation

DGX agent

arXiv:2512.14954v2 Announce Type: replace Abstract: Computing next-token likelihood ratios between two language models (LMs) is a standard task in training paradigms such as knowledge distillation. Si

researcharxiv-cs-cl
7 May 2026
Hardware

Deep Wave Network for Modeling Multi-Scale Physical Dynamics

DGX agent

arXiv:2605.04198v1 Announce Type: new Abstract: Performance of deep learning models is strongly governed by architectural capacity, with width and depth as primary controls. However, in physical-scien

hardwarearxiv-cs-lg
7 May 2026
Model Releases

I don't really ever trust benchmarks, so ocassionally I'm stress-vibe-testing a bunch of new models on some very complex agent work (hundred…

DGX agent

I don't really ever trust benchmarks, so ocassionally I'm stress-vibe-testing a bunch of new models on some very complex agent work (hundreds of tools, not your simple coding agent stuff) and to my su

model-releasesclem-delangue--x
7 May 2026
Research

Lookahead Drifting Model

DGX agent

arXiv:2605.04060v1 Announce Type: cross Abstract: Recently, a new paradigm named drifting model has been proposed for mapping distributions, which achieves the SOTA image generation performance over I

researcharxiv-cs-cv
7 May 2026
Tutorials

Multi Language Models for On-the-Fly Syntax Highlighting

DGX agent

arXiv:2510.04166v2 Announce Type: replace-cross Abstract: Syntax highlighting is a critical feature in modern software development environments, enhancing code readability and developer productivity.

tutorialsarxiv-cs-ai
7 May 2026
Model Releases

New Anthropic research: Natural Language Autoencoders. Models like Claude talk in words but think in numbers. The numbers—called activations…

DGX agent

New Anthropic research: Natural Language Autoencoders. Models like Claude talk in words but think in numbers. The numbers—called activations—encode Claude’s thoughts, but not in a language we can read

model-releasesboris-cherny--x
7 May 2026
Agents

Open Models Make Agentic Batch Processing Economically Viable A lot of world’s work looks like “Do X for EVERY Y” - read every trace - respo…

DGX agent

Open Models Make Agentic Batch Processing Economically Viable A lot of world’s work looks like “Do X for EVERY Y” - read every trace - respond to every email - deep dive into every document - enrich e

agentsharrison-chase--x
7 May 2026
Model Releases

PSK at SemEval-2026 Task 9: Multilingual Polarization Detection Using Ensemble Gemma Models with Synthetic Data Augmentation

DGX agent

arXiv:2605.05159v1 Announce Type: new Abstract: We present our system for SemEval-2026 Task 9: Multilingual Polarization Detection, a binary classification task spanning 22 languages. Our approach fin

model-releasesarxiv-cs-cl
7 May 2026
Research

Right Model, Right Time: Real-Time Cascaded-Fidelity MPC for Bipedal Walking

DGX agent

arXiv:2605.04607v1 Announce Type: new Abstract: This paper presents a multi-phase whole-body model predictive control approach for bipedal walking, combining a detailed whole-body model in the near ho

researcharxiv-cs-ro
7 May 2026
Model Releases

RLearner-LLM: Balancing Logical Grounding and Fluency in Large Language Models via Hybrid Direct Preference Optimization

DGX agent

arXiv:2605.04539v1 Announce Type: new Abstract: Direct Preference Optimization (DPO), the efficient alternative to PPO-based RLHF, falls short on knowledge-intensive generation: standard preference si

model-releasesarxiv-cs-cl
7 May 2026
Applications

So Mythos was, indeed, not marketing hype. Remember this is a general purpose model that just happens to be good at finding exploits because…

DGX agent

So Mythos was, indeed, not marketing hype. Remember this is a general purpose model that just happens to be good at finding exploits because good models are good at lots of things. Expect similar from

applicationsethan-mollick--x
7 May 2026
Model Releases

TCM-Serve: Modality-aware Scheduling for Multimodal Large Language Model Inference

DGX agent

arXiv:2603.26498v2 Announce Type: replace-cross Abstract: Multimodal Large Language Models (MLLMs) power platforms like ChatGPT, Gemini, and Copilot, enabling richer interactions with text, images, an

model-releasesarxiv-cs-ai
7 May 2026
Model Releases

The Shape of Beliefs: Geometry, Dynamics, and Interventions along Representation Manifolds of Language Models' Posteriors

DGX agent

arXiv:2602.02315v2 Announce Type: replace Abstract: Large language models (LLMs) form implicit beliefs (posteriors over latent variables) from prompts, but we lack a mechanistic account of how these b

model-releasesarxiv-cs-cl
7 May 2026
Safety

Uncertainty-Aware Exploratory Direct Preference Optimization for Multimodal Large Language Models

DGX agent

arXiv:2605.04874v1 Announce Type: cross Abstract: Direct Preference Optimization (DPO) has proven to be an effective solution for mitigating hallucination in Multimodal Large Language Models (MLLMs) b

safetyarxiv-cs-cl
7 May 2026
Model Releases

We recently found some instances of CoT grading during the training of previously deployed models after building a system that scans all Ope…

DGX agent

We recently found some instances of CoT grading during the training of previously deployed models after building a system that scans all OpenAI RL runs for accidental CoT grading. We did not find clea

model-releasesopenai--x
7 May 2026
Local Ai

A Few-Step Generative Model on Cumulative Flow Maps

DGX agent

arXiv:2605.03623v1 Announce Type: new Abstract: We propose a unified, few-step generative modeling framework based on cumulative flow maps for long-range transport in probability space, inspired by fl

local-aiarxiv-cs-lg
6 May 2026
Applications

A Unified Framework for Tabular Generative Modeling: Loss Functions, Benchmarks, and Improved Multi-objective Bayesian Optimization Approaches

DGX agent

arXiv:2405.16971v2 Announce Type: replace Abstract: Deep learning (DL) models require extensive data to achieve strong performance and generalization. Deep generative models (DGMs) offer a solution by

applicationsarxiv-cs-lg
6 May 2026
Safety

Geometry Forcing: Marrying Video Diffusion and 3D Representation for Consistent World Modeling

DGX agent

arXiv:2507.07982v2 Announce Type: replace Abstract: Videos inherently represent 2D projections of a dynamic 3D world. However, our analysis suggests that video diffusion models trained solely on raw v

safetyarxiv-cs-cv
6 May 2026
Model Releases

Google Chrome silently installs a ~4GB Gemini Nano model on desktop devices; Google says it has been offered since 2024 and users can remove it via settings (Ben Schoon/9to5Google)

DGX agent

Ben Schoon / 9to5Google: Google Chrome silently installs a ~4GB Gemini Nano model on desktop devices; Google says it has been offered since 2024 and users can remove it via settings — The ongoing marc

model-releasestechmeme
6 May 2026
Safety

GRPO-TTA: Test-Time Visual Tuning for Vision-Language Models via GRPO-Driven Reinforcement Learning

DGX agent

arXiv:2605.03403v1 Announce Type: new Abstract: Group Relative Policy Optimization (GRPO) has recently shown strong performance in post-training large language models and vision-language models. It ra

safetyarxiv-cs-cv
6 May 2026
Model Releases

Hybrid Models for Natural Language Reasoning: The Case of Syllogistic Logic

DGX agent

arXiv:2510.09472v2 Announce Type: replace Abstract: Despite the remarkable progress in neural models, their ability to generalize, a cornerstone for applications such as logical reasoning, remains a c

model-releasesarxiv-cs-cl
6 May 2026
Applications

Joint Relational Database Generation via Graph-Conditional Diffusion Models

DGX agent

arXiv:2505.16527v2 Announce Type: replace Abstract: Building generative models for relational databases (RDBs) is important for many applications, such as privacy-preserving data release and augmentin

applicationsarxiv-cs-lg
6 May 2026
Safety

Khala: Scaling Acoustic Token Language Models Toward High-Fidelity Music Generation

DGX agent

arXiv:2605.01790v1 Announce Type: cross Abstract: A common design pattern in high-quality music generation is to handle structure and fidelity in different representation spaces: a generator first mod

safetyarxiv-cs-ai
6 May 2026
Agents

Latent State Design for World Models under Sufficiency Constraints

DGX agent

arXiv:2605.01694v1 Announce Type: new Abstract: A world model matters to an agent only through the state it constructs. That state must preserve some information, discard other information, and suppor

agentsarxiv-cs-ai
6 May 2026
Model Releases

MHPR: Multidimensional Human Perception and Reasoning Benchmark for Large Vision-Languate Models

DGX agent

arXiv:2605.03485v1 Announce Type: new Abstract: Multidimensional human understanding is essential for real-world applications such as film analysis and virtual digital humans, yet current LVLM benchma

model-releasesarxiv-cs-cv
6 May 2026
Model Releases

MRC is already deployed across all of OpenAI’s largest supercomputers that we use to train frontier models, including our site with @Oracle …

DGX agent

MRC is already deployed across all of OpenAI’s largest supercomputers that we use to train frontier models, including our site with @Oracle Cloud Infrastructure (OCI) in Abilene, Texas, and in @Micros

model-releasesopenai--x
6 May 2026
Tutorials

SCPRM: A Schema-aware Cumulative Process Reward Model for Knowledge Graph Question Answering

DGX agent

arXiv:2605.02819v1 Announce Type: new Abstract: Large language models excel at complex reasoning, yet evaluating their intermediate steps remains challenging. Although process reward models provide st

tutorialsarxiv-cs-ai
6 May 2026
Local Ai

SHIELD: A Diverse Clinical Note Dataset and Distilled Small Language Models for Enterprise-Scale De-identification

DGX agent

arXiv:2605.03301v1 Announce Type: new Abstract: De-identification of clinical text remains essential for secondary use of electronic health records (EHRs), yet public benchmarks such as i2b2 2006/2014

local-aiarxiv-cs-cl
6 May 2026
Local Ai

Tencent is about to release an anime video model (AniMatrix).

DGX agent

Tencent launched Hunyuan Video in December 2024, an open-source AI video generation model with 13 billion parameters that supports text-to-video and image-to-video conversion. The model brings unique

local-air-stablediffusion
6 May 2026
Model Releases

Vibe Code Bench: Evaluating AI Models on End-to-End Web Application Development

DGX agent

arXiv:2603.04601v2 Announce Type: replace-cross Abstract: Code generation has emerged as one of AI's highest-impact use cases, yet existing benchmarks measure isolated tasks rather than the complete '

model-releasesarxiv-cs-cl
6 May 2026
Model Releases

Break the Block: Dynamic-size Reasoning Blocks for Diffusion Large Language Models via Monotonic Entropy Descent with Reinforcement Learning

DGX agent

arXiv:2605.02263v1 Announce Type: new Abstract: Recent diffusion large language models (dLLMs) have demonstrated both effectiveness and efficiency in reasoning via a block-based semi-autoregressive ge

model-releasesarxiv-cs-lg
5 May 2026
Research

Counting as a minimal probe of language model reliability

DGX agent

arXiv:2605.02028v1 Announce Type: new Abstract: Large language models perform strongly on benchmarks in mathematical reasoning, coding and document analysis, suggesting a broad ability to follow instr

researcharxiv-cs-cl
5 May 2026
Model Releases

Empowering Heterogeneous Graph Foundation Models via Decoupled Relation Alignment

DGX agent

arXiv:2605.00731v1 Announce Type: cross Abstract: While Graph Foundation Models (GFMs) have achieved remarkable success in homogeneous graphs, extending them to multi-domain heterogeneous graphs (MDHG

model-releasesarxiv-cs-ai
5 May 2026
Model Releases

GaMMA: Towards Joint Global-Temporal Music Understanding in Large Multimodal Models

DGX agent

arXiv:2605.00371v1 Announce Type: cross Abstract: In this paper, we propose GaMMA, a state-of-the-art (SoTA) large multimodal model (LMM) designed to achieve comprehensive musical content understandin

model-releasesarxiv-cs-ai
5 May 2026
← Previous
1…144145146147148…1262
Next →