AI Wiki
TimelineEvolutionGraphStatusAsk wiki
Live from Git
AI Wiki
TimelineEvolutionGraphStatusAsk wiki
Live from Git
Filter entries
Categories
  • All entries85,188
  • Agents7,322
  • Applications5,231
  • Concepts5
  • Hardware1,770
  • Industry6,109
  • Local Ai4,762
  • Model Releases22,797
  • Research19,333
  • Safety12,893
  • Syntheses17
  • Tools1,670
  • Tutorials3,279

Source
HumanDGX agent

Content type
AllBlogX PostPaperYouTubeRedditGitHub
Clear filters
Categories
  • All entries85,188
  • Agents7,322
  • Applications5,231
  • Concepts5
  • Hardware1,770
  • Industry6,109
  • Local Ai4,762
  • Model Releases22,797
  • Research19,333
  • Safety12,893
  • Syntheses17
  • Tools1,670
  • Tutorials3,279

Source
HumanDGX agent

Content type
85,188Total entries
1Added by human
85,187Found by agent
12Categories

Knowledge catalogue

Search: “model-releases”

GridTimelineEvolution
22,805 results
Model Releases

TravelEval: A Comprehensive Benchmarking Framework for Evaluating LLM-Powered Travel Planning Agents

DGX agent

arXiv:2606.01046v1 Announce Type: new Abstract: The development of Large Language Models (LLMs) has significantly improved travel planning applications, yet evaluating such models is limited by existi

model-releasesarxiv-cs-ai
2 Jun 2026
AllBlogX PostPaperYouTubeRedditGitHub
Clear filters
Model Releases

Tree-Structured Parzen Estimator: Understanding Its Algorithm Components and Their Roles for Better Empirical Performance

DGX agent

arXiv:2304.11127v5 Announce Type: replace-cross Abstract: Recent scientific advances require complex experiment design, necessitating the meticulous tuning of many experiment parameters. Tree-structur

model-releasesarxiv-cs-ai
2 Jun 2026
Model Releases

Trump signs executive order to review AI models before they’re released

DGX agent

President Donald Trump signed an executive order Tuesday creating a 'voluntary framework' for AI companies to share their frontier models with the federal government before they're released 'to promot

model-releasesthe-verge-ai
2 Jun 2026
Model Releases

TrustLDM: Benchmarking Trustworthiness in Language Diffusion Models

DGX agent

arXiv:2606.00023v1 Announce Type: cross Abstract: The rapid development of Language Diffusion Models (LDMs) challenges the dominant position of auto-regressive competitors in language processing. Howe

model-releasesarxiv-cs-ai
2 Jun 2026
Model Releases

Truth, Trust, and Trouble: Medical AI on the Edge

DGX agent

arXiv:2507.02983v3 Announce Type: replace-cross Abstract: Large Language Models (LLMs) hold significant promise for transforming digital health by enabling automated medical question answering. Howeve

model-releasesarxiv-cs-ai
2 Jun 2026
Model Releases

Truthful AI Advisors: A Pre-Specified Benchmark for Large Language Model Honesty Under Preference Misalignment

DGX agent

arXiv:2606.01456v1 Announce Type: cross Abstract: Large language models are increasingly deployed as advisors whose objective is not aligned with the user's: recommenders optimize for engagement, sale

model-releasesarxiv-cs-cl
2 Jun 2026
Model Releases

TukaBench: A Culturally Grounded Jailbreak Benchmark for African Languages

DGX agent

arXiv:2606.01322v1 Announce Type: cross Abstract: Safety evaluation of Large Language Models (LLMs) remains heavily English-centric, leaving Low-Resource Languages (LRLs), particularly African ones, c

model-releasesarxiv-cs-ai
2 Jun 2026
Model Releases

Turning Back Without Forgetting: Selective Backward Refinement for Parameter-Efficient Continual Learning

DGX agent

arXiv:2606.01379v1 Announce Type: new Abstract: While prompt-based parameter-efficient continual learning mitigates catastrophic forgetting by isolating task-specific prompts, this isolation also limi

model-releasesarxiv-cs-lg
2 Jun 2026
Model Releases

TVIR: Building Deep Research Agents Towards Text--Visual Interleaved Report Generation

DGX agent

arXiv:2606.02320v1 Announce Type: new Abstract: Deep Research Agents have shown strong capability in multi-step information retrieval, reasoning, and long-form report generation, but existing benchmar

model-releasesarxiv-cs-cl
2 Jun 2026
Model Releases

Uber says it has limited all employees to $1,500 in monthly token spending per AI coding tool 'to responsibly encourage agentic AI adoption' (Natalie Lung/Bloomberg)

DGX agent

Natalie Lung / Bloomberg: Uber says it has limited all employees to $1,500 in monthly token spending per AI coding tool “to responsibly encourage agentic AI adoption” — Uber Technologies Inc. has set

model-releasestechmeme
2 Jun 2026
Model Releases

Uncovering Competency Gaps in Large Language Models and Their Benchmarks

DGX agent

arXiv:2512.20638v2 Announce Type: replace-cross Abstract: The evaluation of large language models relies heavily on standardized benchmarks. These benchmarks provide useful aggregated metrics, but can

model-releasesarxiv-cs-ai
2 Jun 2026
Model Releases

Understanding Identity Continuity in Thermal Video through Scene-Level Consistency

DGX agent

arXiv:2606.01694v1 Announce Type: cross Abstract: Thermal pedestrian MOT remains challenging because weak appearance cues and frequent detection interruptions cause severe trajectory fragmentation. We

model-releasesarxiv-cs-ai
2 Jun 2026
Model Releases

Understanding LLM Behavior in Multi-Target Cross-Lingual Summarization

DGX agent

arXiv:2606.01252v1 Announce Type: cross Abstract: Multi-target cross-lingual text summarization (MTXLS), which summarizes a source document into multiple target languages, is increasingly important as

model-releasesarxiv-cs-ai
2 Jun 2026
Model Releases

UniD^3: A Knowledge Graph-Enhanced RAG Framework for Drug-Disease Discovery and Reasoning

DGX agent

arXiv:2606.01394v1 Announce Type: new Abstract: Systematic characterization of drug-disease relationships is essential for drug discovery and repurposing, yet is hindered by the heterogeneity and rapi

model-releasesarxiv-cs-cl
2 Jun 2026
Model Releases

Universal Quantum Transformer

DGX agent

arXiv:2606.00045v1 Announce Type: new Abstract: Classical continuous-space neural networks fundamentally struggle to lock into exact mathematical symmetries, such as modular arithmetic and non-commuta

model-releasesarxiv-cs-ai
2 Jun 2026
Model Releases

Unlocking the Black Box of Latent Reasoning: An Interpretability-Guided Approach to Intervention

DGX agent

arXiv:2606.01243v1 Announce Type: new Abstract: Latent reasoning enables Large Language Models (LLMs) to perform multi-step inference within continuous hidden states, offering efficiency gains over ex

model-releasesarxiv-cs-cl
2 Jun 2026
Model Releases

URDF-Anything+: End-to-End Generation for Simulation-Ready Articulated Assets

DGX agent

arXiv:2603.14010v2 Announce Type: replace Abstract: Articulated objects are fundamental for robotics, simulation of physics, and interactive virtual environments. However, recovering them from visual

model-releasesarxiv-cs-ro
2 Jun 2026
Model Releases

v-HUB: A Benchmark for Video Humor Understanding from Vision and Sound

DGX agent

arXiv:2509.25773v3 Announce Type: replace-cross Abstract: AI models capable of comprehending humor hold real-world promise -- for example, enhancing engagement in human-machine interactions. To gauge

model-releasesarxiv-cs-ai
2 Jun 2026
Model Releases

v0.30.1: llm: ignore llama-server SSE ping comments (#16443)

DGX agent

Ollama v0.30.1 addresses an issue where the LLM component now ignores Server-Sent Events (SSE) ping comments from llama-server, resolving problem #16443. This fix improves the stability and reliabilit

model-releasesollama-releases
2 Jun 2026
Model Releases

Value Flows

DGX agent

arXiv:2510.07650v4 Announce Type: replace-cross Abstract: While most reinforcement learning methods today flatten the distribution of future returns to a single scalar value, distributional RL methods

model-releasesarxiv-cs-ai
2 Jun 2026
Model Releases

VESTA: Visual Exploration with Statistical Tool Agents

DGX agent

arXiv:2606.00384v1 Announce Type: new Abstract: Fitting quantitative models to data is a central step in scientific workflows, yet it remains one of the least automated. Recent agent-based systems lev

model-releasesarxiv-cs-ai
2 Jun 2026
Model Releases

Vision-language Models for Driver Monitoring Systems: A Driver Activity Description Dataset

DGX agent

arXiv:2606.02273v1 Announce Type: new Abstract: Understanding subtle driver actions is essential for building reliable driver monitoring systems. Existing visionlanguage models (VLMs) are trained on g

model-releasesarxiv-cs-cv
2 Jun 2026
Model Releases

Visual-Noise Guided In-Context Distillation for Multimodal Large Language Model Unlearning

DGX agent

arXiv:2606.00105v1 Announce Type: cross Abstract: Multimodal Large Language Models (MLLMs) have achieved remarkable progress on vision-language tasks, but they may also memorize and expose sensitive o

model-releasesarxiv-cs-ai
2 Jun 2026
Model Releases

VLBM: Variational Latent Basis Modeling for OOD Robust Multivariate Time Series Forecasting

DGX agent

arXiv:2606.02138v1 Announce Type: cross Abstract: Out of distribution (OOD) events in multivariate time series forecasting are rare but often dominate real world risk, making average case forecasting

model-releasesarxiv-cs-ai
2 Jun 2026
Model Releases

VocSim: A Training-free Benchmark for Zero-shot Content Identity in Single-source Audio

DGX agent

arXiv:2512.10120v2 Announce Type: replace-cross Abstract: General-purpose audio representations aim to map acoustically variable instances of the same event to nearby points, resolving content identit

model-releasesarxiv-cs-ai
2 Jun 2026
Model Releases

WAON: A Large-Scale Japanese Image-Text Dataset for Cultural Adaptation in Contrastive Vision-Language Models

DGX agent

arXiv:2510.22276v3 Announce Type: replace-cross Abstract: Contrastive vision-language models have achieved remarkable progress through large-scale pretraining. Recent work has shown that removing Engl

model-releasesarxiv-cs-cl
2 Jun 2026
Model Releases

We wrapped a live session on M3 yesterday with the @togethercompute team & our researchers @zpysky1125 and @HaohaiSun A few highlights 🧵 1.…

DGX agent

We wrapped a live session on M3 yesterday with the @togethercompute team & our researchers @zpysky1125 and @HaohaiSun A few highlights 🧵 1. MSA (MiniMax Sparse Attention) is the star ⭐️. Unlike CSA/HC

model-releasestogether-ai--x
2 Jun 2026
Model Releases

We're sponsoring a hackathon to scale down. Hosted by our friends @huggingface and @Gradio, we want working with models to feel like yours a…

DGX agent

We're sponsoring a hackathon to scale down. Hosted by our friends @huggingface and @Gradio, we want working with models to feel like yours again. Small enough that it's inexpensive to run, big enough

model-releasescohere--x
2 Jun 2026
Model Releases

What Do LLMs Know About Alzheimer's Disease? Multi-loss Fine-Tuning and Probing for AD Detection

DGX agent

arXiv:2602.11177v2 Announce Type: replace-cross Abstract: Reliable early detection of Alzheimer's disease (AD) is challenging, particularly due to the limited availability of labeled data. While large

model-releasesarxiv-cs-ai
2 Jun 2026
Model Releases

What to Format and How: A Benchmark and Workflow Approach for Document Formatting

DGX agent

arXiv:2606.01936v1 Announce Type: new Abstract: Recent advances in large language models (LLMs) have opened up new possibilities for automated document formatting. However, real-world formatting often

model-releasesarxiv-cs-cl
2 Jun 2026
Model Releases

When AI Benchmarks Plateau: A Systematic Study of Benchmark Saturation

DGX agent

arXiv:2602.16763v2 Announce Type: replace Abstract: Artificial intelligence benchmarks are an important mechanism for measuring model progress and guiding deployment decisions. However, benchmarks qui

model-releasesarxiv-cs-ai
2 Jun 2026
Model Releases

When and How Much to Imagine: Adaptive Test-Time Scaling with World Models for Visual Spatial Reasoning

DGX agent

arXiv:2602.08236v2 Announce Type: replace-cross Abstract: Despite rapid progress in MLLMs, visual spatial reasoning remains unreliable when correct answers depend on how a scene would appear under uns

model-releasesarxiv-cs-ai
2 Jun 2026
Model Releases

When Jokes Cross the Line: Analyzing Regular Humor and Dark Humor in YouTube Shorts

DGX agent

arXiv:2606.00046v1 Announce Type: cross Abstract: Video platforms such as YouTube have reshaped how users engage with entertainment and information, emphasizing brief, highly engaging content such as

model-releasesarxiv-cs-ai
2 Jun 2026
Model Releases

When Parallelism Pays Off: Cohesion-Aware Task Partitioning for Multi-Agent Coding

DGX agent

arXiv:2606.00953v1 Announce Type: new Abstract: Multi-agent Large Language Model (LLM) systems offer a way to decompose complex tasks, such as coding, through parallelization and context isolation. Ho

model-releasesarxiv-cs-lg
2 Jun 2026
Model Releases

When Safe Skills Collide: Measuring Compositional Risk in Agent Skill Ecosystems

DGX agent

arXiv:2606.00448v1 Announce Type: cross Abstract: LLM agents increasingly rely on community-contributed skills that expand an agent's operational capability set. We study a core safety problem in agen

model-releasesarxiv-cs-ai
2 Jun 2026
Model Releases

When Single Answer Is Not Enough: Rethinking Single-Step Retrosynthesis Benchmarks for LLMs

DGX agent

arXiv:2602.03554v2 Announce Type: replace-cross Abstract: Recent progress has expanded the use of large language models (LLMs) in drug discovery, including synthesis planning. However, objective evalu

model-releasesarxiv-cs-ai
2 Jun 2026
Model Releases

Where Do Deep-Research Agents Go Wrong? Span-Level Error Localization in Agent Trajectories

DGX agent

arXiv:2606.02060v1 Announce Type: new Abstract: Deep-research agents solve tasks through long trajectories of search, tool use, evidence inspection, and answer synthesis. Evaluation based on final ans

model-releasesarxiv-cs-ai
2 Jun 2026
Model Releases

Where to Look: Can Foundation Models Reach a Target Viewpoint Through Active Exploration?

DGX agent

arXiv:2606.01247v1 Announce Type: new Abstract: Humans can reproduce the viewpoint specified by a target image through active head and body motion, yet spatial intelligence in foundation models has la

model-releasesarxiv-cs-cv
2 Jun 2026
Model Releases

Which Leakage Types Matter? A Quantitative Landscape Across 2,047 Benchmark Datasets

DGX agent

arXiv:2604.04199v2 Announce Type: replace Abstract: Twenty-eight within-subject counterfactual experiments across 2,047 iid tabular datasets, plus a boundary experiment on 129 temporal datasets, measu

model-releasesarxiv-cs-lg
2 Jun 2026
Model Releases

Why Not Hyperparameter-Friendly Optimisation? A Monotonic Adaptive Norm Rescaling Approach For Long-Tailed Recognition

DGX agent

arXiv:2606.02526v1 Announce Type: cross Abstract: Long-tailed recognition poses a significant challenge for deep learning. The two-stage decoupling paradigm, which separates representation learning fr

model-releasesarxiv-cs-ai
2 Jun 2026
Model Releases

WildCat: Near-Linear Attention in Theory and Practice

DGX agent

arXiv:2602.10056v2 Announce Type: replace Abstract: We introduce WildCat, a high-accuracy, low-cost approach to compressing the attention mechanism in neural networks. While attention is a staple of m

model-releasesarxiv-cs-lg
2 Jun 2026
Model Releases

Wordle 1,808 5/6 ⬛⬛⬛⬛⬛ 🟨🟨⬛⬛⬛ ⬛🟨🟩🟨⬛ 🟩🟩🟩🟩⬛ 🟩🟩🟩🟩🟩

DGX agent

This post shows the solution path for Wordle puzzle #1,808, solved in 5 guesses, with a visual representation of each guess's results using colored tiles indicating correct letters (green), misplaced

model-releasesanthropic--x
2 Jun 2026
Model Releases

Workflows are the biggest upgrade to Claude Code’s capabilities since skills and subagents. I dove deep into it with @sidbid to figure out b…

DGX agent

Workflows are the biggest upgrade to Claude Code’s capabilities since skills and subagents. I dove deep into it with @sidbid to figure out best practices, examples and more. I’m particularly excited a

model-releasesthariq--x
2 Jun 2026
Model Releases

WorldCache: Accelerating World Models for Free via Heterogeneous Token Caching

DGX agent

arXiv:2603.06331v2 Announce Type: replace Abstract: Diffusion-based world models have shown strong potential for unified world simulation, but the iterative denoising remains too costly for interactiv

model-releasesarxiv-cs-cv
2 Jun 2026
Model Releases

WorldCoder-Bench: Benchmarking Physically Grounded 3D World Synthesis

DGX agent

arXiv:2606.01869v1 Announce Type: new Abstract: Large language models (LLMs) are increasingly asked not only to write static interfaces, but to construct executable interactive worlds from natural lan

model-releasesarxiv-cs-ai
2 Jun 2026
Model Releases

WorldLens: Full-Spectrum Evaluations of Driving World Models in Real World

DGX agent

arXiv:2512.10958v2 Announce Type: replace Abstract: Generative world models are reshaping embodied AI, enabling agents to synthesize realistic 4D driving environments that look convincing but often fa

model-releasesarxiv-cs-cv
2 Jun 2026
Model Releases

Worlds Within Words: Translating Culture in Ancient Chinese Texts with Multi-Agent Coordination

DGX agent

arXiv:2606.01276v1 Announce Type: new Abstract: Large language model (LLM)-based machine translation has advanced cross-cultural communication, yet it still struggles with culture-loaded words (CLWs)

model-releasesarxiv-cs-cl
2 Jun 2026
Model Releases

WUSH: Near-Optimal Adaptive Transforms for LLM Quantization

DGX agent

arXiv:2512.00956v3 Announce Type: replace Abstract: Quantizing LLM weights and activations is a standard approach for efficient deployment, but a few extreme outliers can stretch the dynamic range and

model-releasesarxiv-cs-lg
2 Jun 2026
← Previous
1…237238239240241…476
Next →