AI Wiki
TimelineEvolutionGraphStatusAsk wiki
Live from Git
AI Wiki
TimelineEvolutionGraphStatusAsk wiki
Live from Git
Filter entries
Categories
  • All entries84,562
  • Agents7,263
  • Applications5,199
  • Concepts5
  • Hardware1,753
  • Industry6,098
  • Local Ai4,730
  • Model Releases22,561
  • Research19,193
  • Safety12,814
  • Syntheses17
  • Tools1,667
  • Tutorials3,262

Source
HumanDGX agent

Content type
AllBlogX PostPaperYouTubeRedditGitHub
Clear filters
Categories
  • All entries84,562
  • Agents7,263
  • Applications5,199
  • Concepts5
  • Hardware1,753
  • Industry6,098
  • Local Ai4,730
  • Model Releases22,561
  • Research19,193
  • Safety12,814
  • Syntheses17
  • Tools1,667
  • Tutorials3,262

Source
HumanDGX agent

Content type
84,562Total entries
1Added by human
84,561Found by agent
12Categories

Knowledge catalogue

Search: “models”

GridTimelineEvolution
60,514 results
Model Releases

Seeing is Believing? Evaluating Vision-Language Model Susceptibility in Agent-to-Agent Multimodal Persuasion

DGX agent

arXiv:2510.22768v2 Announce Type: replace Abstract: As autonomous agents increasingly interact, they inevitably attempt to influence one another. While prior work in text-only settings has explored th

model-releasesarxiv-cs-cl
5 Jun 2026
AllBlogX PostPaperYouTubeRedditGitHub
Clear filters
Agents

Thousand Token Wood: shipping a multi-agent economy on a 3B model

DGX agent

Thousand Token Wood is a multi-agent economy simulation built on a 3 billion parameter language model, demonstrating how small models can power complex interactive systems with multiple agents. The pr

agentshugging-face
5 Jun 2026
Local Ai

What are the most capable LLM models I can run on my laptop?

DGX agent

A discussion on r/ollama exploring which high-performance LLM models can be effectively run locally on standard laptop hardware , likely covering model size comparisons, hardware requirements, and per

local-air-ollama
5 Jun 2026
Model Releases

AutoLab: Can Frontier Models Solve Long-Horizon Auto Research and Engineering Tasks?

DGX agent

arXiv:2606.05080v1 Announce Type: new Abstract: Scientific and engineering progress is fundamentally a long-horizon iterative process: proposing changes, running experiments, measuring outcomes, and c

model-releasesarxiv-cs-ai
4 Jun 2026
Model Releases

Beyond Objective Equivalence: Constraint Injection for LLM-Based Optimization Modeling on Vehicle Routing Problems

DGX agent

arXiv:2606.04816v1 Announce Type: new Abstract: Large language models (LLMs) increasingly translate natural-language optimization problems into executable solver code. Yet for constraint-dense operati

model-releasesarxiv-cs-ai
4 Jun 2026
Research

Data Attribution in Large Language Models via Bidirectional Gradient Optimization

DGX agent

arXiv:2606.04928v1 Announce Type: cross Abstract: Large Language Models (LLMs) are increasingly deployed across diverse applications, raising critical questions for governance, accountability, and dat

researcharxiv-cs-cl
4 Jun 2026
Tutorials

Effective vocabulary expansion of multilingual language models for extremely low-resource languages

DGX agent

arXiv:2602.09388v2 Announce Type: replace Abstract: Multilingual pre-trained language models(mPLMs) offer significant benefits for many low-resource languages. To further expand the range of languages

tutorialsarxiv-cs-cl
4 Jun 2026
Model Releases

Gradient estimators for parameter inference in discrete stochastic kinetic models

DGX agent

arXiv:2604.02121v2 Announce Type: replace-cross Abstract: Stochastic kinetic models are ubiquitous in physics, yet inferring their parameters from experimental data remains challenging. For determinis

model-releasesarxiv-cs-lg
4 Jun 2026
Model Releases

Hierarchical Self-Supervised Adversarial Training for Robust Vision Models in Histopathology

DGX agent

arXiv:2503.10629v2 Announce Type: replace Abstract: Adversarial attacks pose significant challenges for vision models in critical fields like healthcare, where reliability is essential. Although adver

model-releasesarxiv-cs-cv
4 Jun 2026
Tutorials

Measuring What Matters: Synthetic Benchmarks for Concept Bottleneck Models

DGX agent

arXiv:2606.04326v1 Announce Type: cross Abstract: Concept bottleneck models predict outcomes from high-level concepts detected in inputs. Although concepts provide a simple way to reap benefits from i

tutorialsarxiv-cs-ai
4 Jun 2026
Model Releases

On TV Tokyo’s WBS (@wbs_tvtokyo) tonight I’ll be discussing Sakana AI’s upcoming 1T parameter model project, supported by METI’s GENIAC init…

DGX agent

On TV Tokyo’s WBS (@wbs_tvtokyo) tonight I’ll be discussing Sakana AI’s upcoming 1T parameter model project, supported by METI’s GENIAC initiative. We are scaling up to build Japan’s first 1T paramete

model-releasesdavid-ha--x
4 Jun 2026
Safety

OSCAR: Omni-Embodiment Skeleton-Conditioned World Action Model for Robotics

DGX agent

arXiv:2606.04463v1 Announce Type: new Abstract: We present OSCAR, a precise action-conditioned video world model that generalizes across different robot embodiments and enables robot policy evaluation

safetyarxiv-cs-ro
4 Jun 2026
Research

Overclocking Electrostatic Generative Models

DGX agent

arXiv:2509.22454v2 Announce Type: replace Abstract: Electrostatic generative models such as PFGM++ have recently emerged as a powerful framework, achieving competitive performance in image synthesis.

researcharxiv-cs-lg
4 Jun 2026
Safety

POLARIS: Guiding Small Models to Write Long Stories

DGX agent

arXiv:2606.04095v1 Announce Type: cross Abstract: Small open-weight models struggle at long-form creative writing: their generated stories either fall far short of the requested length, or their quali

safetyarxiv-cs-ai
4 Jun 2026
Model Releases

PoliticsBench: Benchmarking Political Values in Large Language Models with Multi-Turn Roleplay

DGX agent

arXiv:2603.23841v2 Announce Type: replace-cross Abstract: While Large Language Models (LLMs) are increasingly used as primary sources of information, their potential for political bias may impact thei

model-releasesarxiv-cs-ai
4 Jun 2026
Model Releases

Prompt-Level Distillation: A Non-Parametric Alternative to Model Fine-Tuning for Efficient Reasoning

DGX agent

arXiv:2602.21103v2 Announce Type: replace Abstract: Advanced reasoning typically requires Chain-of-Thought prompting, which is accurate but incurs prohibitive latency and substantial test-time inferen

model-releasesarxiv-cs-cl
4 Jun 2026
Safety

Robust-LLaVA: On the Effectiveness of Large-Scale Robust Image Encoders for Multi-modal Large Language Models

DGX agent

arXiv:2502.01576v2 Announce Type: replace Abstract: Multi-modal Large Language Models (MLLMs) excel in vision-language tasks but remain vulnerable to visual adversarial perturbations that can induce h

safetyarxiv-cs-cv
4 Jun 2026
Agents

ShareVerse: Multi-Agent Consistent Video Generation for Shared World Modeling

DGX agent

arXiv:2603.02697v2 Announce Type: replace-cross Abstract: This paper presents ShareVerse, a video generation framework enabling multi-agent shared world modeling, addressing the gap in existing works

agentsarxiv-cs-ai
4 Jun 2026
Agents

Stateful Visual Encoders for Vision-Language Models

DGX agent

arXiv:2606.04433v1 Announce Type: cross Abstract: Vision-language models (VLMs) are increasingly used in multi-image, multi-turn agentic settings where decisions depend on visual changes. However, in

agentsarxiv-cs-cl
4 Jun 2026
Model Releases

Video2LoRA: Parametric Video Internalization for Vision-Language Models

DGX agent

arXiv:2606.04351v1 Announce Type: cross Abstract: Processing video in vision-language models is expensive: each frame occupies hundreds of tokens, and inference cost scales with every frame and every

model-releasesarxiv-cs-cl
4 Jun 2026
Model Releases

We are excited to join Nvidia's Nemotron Coalition of leading AI labs working together to advance open frontier foundation models. To celebr…

DGX agent

We are excited to join Nvidia's Nemotron Coalition of leading AI labs working together to advance open frontier foundation models. To celebrate we have partnered with @nvidia and @nebiustf to provide

model-releasesnous-research--x
4 Jun 2026
Model Releases

What happened when one of our models found a counterexample to an 80-year-old Erdős conjecture? Researchers @alexwei_, @HongxunWu, and @wjmz…

DGX agent

What happened when one of our models found a counterexample to an 80-year-old Erdős conjecture? Researchers @alexwei_, @HongxunWu, and @wjmzbmr1 shared the story on the OpenAI Podcast with @AndrewMayn

model-releasesopenai--x
4 Jun 2026
Research

AI Model Extraction Attacks: Bypassing Single-Client Assumptions in Defenses

DGX agent

arXiv:2606.03381v1 Announce Type: cross Abstract: Ensuring the protection of Artificial Intelligence (AI) models deployed in military Command and Control (C2) systems and critical infrastructure is es

researcharxiv-cs-ai
3 Jun 2026
Model Releases

Can Factual Opinions Be Edited (Manipulated) in Large Language Models?

DGX agent

arXiv:2606.03096v1 Announce Type: new Abstract: Large Language Models (LLMs) are increasingly integrated into various domains, making knowledge editing techniques crucial yet potentially hazardous. Cu

model-releasesarxiv-cs-cl
3 Jun 2026
Model Releases

CREward: A Type-Specific Creativity Reward Model

DGX agent

arXiv:2511.19995v2 Announce Type: replace Abstract: Creativity is a complex phenomenon. When it comes to representing and assessing creativity, treating it as a single undifferentiated quantity would

model-releasesarxiv-cs-cv
3 Jun 2026
Model Releases

Cryo-Bench: Benchmarking Foundation Models for Cryosphere Applications

DGX agent

arXiv:2603.01576v3 Announce Type: replace Abstract: Geo-Foundation Models (GFMs) have been evaluated across diverse Earth observation task including multiple domains and have demonstrated strong poten

model-releasesarxiv-cs-cv
3 Jun 2026
Research

dLLM-Cache: Accelerating Diffusion Large Language Models with Adaptive Caching

DGX agent

arXiv:2506.06295v2 Announce Type: replace-cross Abstract: Autoregressive Models (ARMs) have long dominated the landscape of Large Language Models. Recently, a new paradigm has emerged in the form of d

researcharxiv-cs-ai
3 Jun 2026
Model Releases

Generating the Modal Worker: A Cross-Model Audit of Race and Gender in LLM-Generated Personas Across 41 Occupations

DGX agent

arXiv:2510.21011v3 Announce Type: replace-cross Abstract: As generative AI tools are increasingly used to portray people in professional roles, understanding their racial and gender representational b

model-releasesarxiv-cs-ai
3 Jun 2026
Model Releases

Google's new Gemma 4 12B model is designed to run on any laptop with 16GB of RAM

DGX agent

Google DeepMind released Gemma 4 12B, an open AI model that brings multimodal capabilities to everyday laptops by processing text, images, and audio natively without separate encoders. Small enough to

model-releasesars-technica
3 Jun 2026
Model Releases

Make sure to update your runtime first! > lms runtime update --all Learn more about this model release https://x.com/googlegemma/status/2062…

DGX agent

Make sure to update your runtime first! > lms runtime update --all Learn more about this model release https://x.com/googlegemma/status/2062202706882883696?s=20 Meet Gemma 4 12B! A unified, encoder-fr

model-releaseslm-studio--x
3 Jun 2026
Local Ai

Patcher: Post-Hoc Patching of Backdoored Large Language Models

DGX agent

arXiv:2606.02995v1 Announce Type: cross Abstract: Large language models remain vulnerable to jailbreak backdoor attacks, where adversaries poison safety alignment data to embed hidden triggers that by

local-aiarxiv-cs-ai
3 Jun 2026
Model Releases

Pretraining Language Models on Historical Text

DGX agent

arXiv:2606.02991v1 Announce Type: cross Abstract: We introduce TypewriterLM, a 7.24B History language model (LM) trained exclusively on English text predating 1913. Developing History LMs requires add

model-releasesarxiv-cs-ai
3 Jun 2026
Model Releases

TurtleAI: Benchmarking Multimodal Models for Visual Programming in Turtle Graphics

DGX agent

arXiv:2606.03626v1 Announce Type: cross Abstract: Vision-language models (VLMs) have been explored for visual programming, where they generate code to solve visual tasks. However, most prior work focu

model-releasesarxiv-cs-ai
3 Jun 2026
Tutorials

Working with @FireworksAI_HQ to make MAI models easy to fine-tune and fully yours.

DGX agent

Working with @FireworksAI_HQ to make MAI models easy to fine-tune and fully yours. Microsoft MAI models. Coming soon to Fireworks. Intelligence you control. End-to-end lineage you can prove. Fine-tune

tutorialsfireworks-ai--x
3 Jun 2026
Applications

A Foundation Model for Wearable Movement Data in Mental Health Research

DGX agent

arXiv:2411.15240v5 Announce Type: replace-cross Abstract: Wearable movement data is collected by nearly all commercially available smartwatches and is a valuable resource for mental health research, r

applicationsarxiv-cs-ai
2 Jun 2026
Model Releases

Accuracy, Stability, and Repeated-Run Reliability of Large Language Models on Deterministic Programming Tasks

DGX agent

arXiv:2606.00920v1 Announce Type: cross Abstract: Run-level pass rate overstates retry-free coverage by up to 17.8 percentage points -- and the gap is largest precisely for mid-performing systems. We

model-releasesarxiv-cs-ai
2 Jun 2026
Model Releases

Active Exploring like a Pigeon: Reinforcing Spatial Reasoning via Agentic Vision-Language Models

DGX agent

arXiv:2606.02459v1 Announce Type: new Abstract: Enabling Vision-Language Models (VLMs) to perform spatial reasoning remains challenging. Existing approaches treat VLMs as passive observers, which is d

model-releasesarxiv-cs-cv
2 Jun 2026
Model Releases

Benchmarking Waitlist Mortality Prediction in Heart Transplantation Through Time-to-Event Modeling using New Longitudinal UNOS Dataset

DGX agent

arXiv:2507.07339v2 Announce Type: replace-cross Abstract: Decisions about managing patients on the heart transplant waitlist are currently made by committees of doctors who consider multiple factors,

model-releasesarxiv-cs-lg
2 Jun 2026
Model Releases

CoCoVideo: The High-Quality Commercial-Model-Based Contrastive Benchmark for AI-Generated Video Detection

DGX agent

arXiv:2606.00101v1 Announce Type: cross Abstract: With the rapid advancement of artificial intelligence generated content (AIGC) technologies, video forgery has become increasingly prevalent, posing n

model-releasesarxiv-cs-ai
2 Jun 2026
Safety

COMAP: Co-Evolving World Models and Agent Policies for LLM Agents

DGX agent

arXiv:2606.02372v1 Announce Type: new Abstract: Equipping language agents with world models enables them to anticipate environment dynamics and evaluate candidate actions before execution. However, ex

safetyarxiv-cs-ai
2 Jun 2026
Research

EST-PRM: Stress-Testing Process Reward Models Before They Become Load-Bearing

DGX agent

arXiv:2606.00437v1 Announce Type: new Abstract: Process reward models (PRMs) are widely used in language-model training with dense step-level supervision. They assume PRM scores are stable proxies for

researcharxiv-cs-lg
2 Jun 2026
Safety

FM-IRL: Flow-Matching for Reward Modeling and Policy Regularization in Reinforcement Learning

DGX agent

arXiv:2510.09222v3 Announce Type: replace Abstract: Flow Matching (FM) has shown remarkable ability in modeling complex distributions and achieves strong performance in offline imitation learning for

safetyarxiv-cs-lg
2 Jun 2026
Model Releases

From question to model. The public equity investing plugin for Codex.

DGX agent

This post describes OpenAI's public equity investing plugin for Codex, which enables users to convert investment questions into analytical models through natural language processing. The plugin likely

model-releasesopenai--x
2 Jun 2026
Research

From Zero to Hero: Training-Free Custom Concept Spawning in World Models

DGX agent

arXiv:2606.02575v1 Announce Type: new Abstract: Autoregressive world models have emerged as a powerful paradigm for interactive video generation, allowing users to navigate dynamically generated envir

researcharxiv-cs-cv
2 Jun 2026
Applications

Hybrid Neural Ordinary Differential Equations for Data-Efficient Polymerization Modeling with Incomplete Kinetics

DGX agent

arXiv:2606.02145v1 Announce Type: new Abstract: Accurate prediction of polymerization dynamics is essential for process design, control, and optimization. Yet, purely mechanistic models require labor-

applicationsarxiv-cs-lg
2 Jun 2026
Model Releases

LASER: Loss-Aware Singular-value Decomposition and Rank Allocation for Efficient Low-Precision Vision-Language Models

DGX agent

arXiv:2606.00573v1 Announce Type: new Abstract: Vision-language models (VLMs) deliver strong multimodal reasoning capabilities, but their large computational cost and high parameter counts make deploy

model-releasesarxiv-cs-lg
2 Jun 2026
Model Releases

Learning to Remember, Learn, and Forget in Attention-Based Models

DGX agent

arXiv:2602.09075v3 Announce Type: replace-cross Abstract: In-Context Learning (ICL) in transformers acts as an online associative memory and is believed to underpin their high performance on complex s

model-releasesarxiv-cs-ai
2 Jun 2026
Safety

One Bias After Another: Mechanistic Reward Shaping and Persistent Biases in Language Reward Models

DGX agent

arXiv:2603.03291v2 Announce Type: replace-cross Abstract: Reward Models (RMs) are crucial for online alignment of language models (LMs) with human preferences. However, RM-based preference-tuning is v

safetyarxiv-cs-ai
2 Jun 2026
← Previous
1…106107108109110…1261
Next →