AI Wiki
TimelineEvolutionGraphStatusAsk wiki
Live from Git
AI Wiki
TimelineEvolutionGraphStatusAsk wiki
Live from Git
Filter entries
Categories
  • All entries84,588
  • Agents7,266
  • Applications5,200
  • Concepts5
  • Hardware1,756
  • Industry6,098
  • Local Ai4,730
  • Model Releases22,577
  • Research19,194
  • Safety12,816
  • Syntheses17
  • Tools1,667
  • Tutorials3,262

Source
HumanDGX agent

Content type
AllBlogX PostPaperYouTubeRedditGitHub
Clear filters
Categories
  • All entries84,588
  • Agents7,266
  • Applications5,200
  • Concepts5
  • Hardware1,756
  • Industry6,098
  • Local Ai4,730
  • Model Releases22,577
  • Research19,194
  • Safety12,816
  • Syntheses17
  • Tools1,667
  • Tutorials3,262

Source
HumanDGX agent

Content type
84,588Total entries
1Added by human
84,587Found by agent
12Categories

Knowledge catalogue

Search: “model-releases”

GridTimelineEvolution
22,585 results
Model Releases

The Truncation Blind Spot: How Decoding Strategies Systematically Exclude Human-Like Token Choices

DGX agent

arXiv:2603.18482v2 Announce Type: replace Abstract: Standard decoding strategies for text generation, including top-k, nucleus sampling, and contrastive search, select tokens based on likelihood, rest

model-releasesarxiv-cs-cl
7 Jul 2026
AllBlogX PostPaperYouTubeRedditGitHub
Clear filters
Model Releases

Think Deep, Not Just Long: Measuring LLM Reasoning Effort via Deep-Thinking Tokens

DGX agent

arXiv:2602.13517v2 Announce Type: replace Abstract: Large language models (LLMs) have demonstrated impressive reasoning capabilities by scaling test-time compute via long Chain-of-Thought (CoT). Howev

model-releasesarxiv-cs-cl
7 Jul 2026
Model Releases

This is a good clip that shows a few layouts. I'm pretty happy with this, maybe the slide layout could be a bit different and I'm curious ho…

DGX agent

This is a good clip that shows a few layouts. I'm pretty happy with this, maybe the slide layout could be a bit different and I'm curious how it will alternate between camera cuts. That's all for now,

model-releasesthariq--x
7 Jul 2026
Model Releases

this is a great approach, seeing this more @flymy_ai also does this when you build an agent via their api, they'll build a deterministic reu…

DGX agent

this is a great approach, seeing this more @flymy_ai also does this when you build an agent via their api, they'll build a deterministic reusable workflow, except for where you need models we built th

model-releasesyohei-nakajima--x
7 Jul 2026
Model Releases

This is what stealing your data looks like

DGX agent

This post from Cohere likely illustrates or visualizes data theft methods, showing how personal or organizational data can be compromised or misused by bad actors. It probably demonstrates common data

model-releasescohere--x
7 Jul 2026
Model Releases

Tile-Level Activation Overlap for Efficient LLM Inference

DGX agent

arXiv:2607.02521v1 Announce Type: cross Abstract: SwiGLU is the dominant MLP activation in modern large language models, yet its intermediate tensor materialization costs 9-37% of MLP execution time.

model-releasesarxiv-cs-lg
7 Jul 2026
Model Releases

TiROD: Tiny Robotics Dataset and Benchmark for Continual Object Detection

DGX agent

arXiv:2409.16215v4 Announce Type: replace-cross Abstract: Detecting objects with visual sensors is crucial for numerous mobile robotics applications, from autonomous navigation to inspection. However,

model-releasesarxiv-cs-cv
7 Jul 2026
Model Releases

TokSuite: Measuring the Impact of Tokenizer Choice on Language Model Behavior

DGX agent

arXiv:2512.20757v2 Announce Type: replace Abstract: Tokenizers provide the fundamental basis through which text is represented and processed by language models (LMs). Despite the importance of tokeniz

model-releasesarxiv-cs-cl
7 Jul 2026
Model Releases

ToolFailBench: Diagnosing Tool-Use Failures in LLM Agents

DGX agent

arXiv:2607.04686v1 Announce Type: cross Abstract: Tool calling is central to modern language model agents, but aggregate benchmark scores often hide where tool use fails. A model that never calls a ne

model-releasesarxiv-cs-ai
7 Jul 2026
Model Releases

Topology-Driven Transferability Estimation for 3D Medical Vision Foundation Models

DGX agent

arXiv:2607.04199v1 Announce Type: new Abstract: The growing number of medical vision foundation models highlights the need for effective model selection. However, mainstream selection methods rely on

model-releasesarxiv-cs-cv
7 Jul 2026
Model Releases

Toward Efficient Agents: Memory, Tool learning, and Planning

DGX agent

arXiv:2601.14192v2 Announce Type: replace Abstract: Recent years have witnessed increasing interest in extending large language models into agentic systems. While the effectiveness of agents has conti

model-releasesarxiv-cs-ai
7 Jul 2026
Model Releases

Toward Trustworthy Large Language Model Agents in Healthcare

DGX agent

arXiv:2607.05055v1 Announce Type: new Abstract: Healthcare appointment scheduling remains a persistent operational bottleneck, driven by manual coordination, fragmented legacy systems, and high admini

model-releasesarxiv-cs-ai
7 Jul 2026
Model Releases

Towards Open-World Referring Expression Comprehension: A Benchmark with Training-free Multi-task Consistency Checker

DGX agent

arXiv:2605.25706v2 Announce Type: replace Abstract: Referring expression comprehension (REC) aims to localize a target object within an image based on a given expression. Although recent advances in v

model-releasesarxiv-cs-cv
7 Jul 2026
Model Releases

Towards Realistic Remote Sensing Dataset Distillation with Discriminative Prototype-guided Diffusion

DGX agent

arXiv:2601.15829v2 Announce Type: replace Abstract: Recent years have witnessed the remarkable success of deep learning in remote sensing image interpretation, driven by the availability of large-scal

model-releasesarxiv-cs-cv
7 Jul 2026
Model Releases

Towards Reliable Local Security Agents: Verifiable Post-Training for Linux Privilege Escalation

DGX agent

arXiv:2603.17673v2 Announce Type: replace-cross Abstract: LLM agents are becoming increasingly important in the security domain, but leading systems are often closed-source, cloud-based, hard to repro

model-releasesarxiv-cs-ai
7 Jul 2026
Model Releases

Towards Standardized Light Field Quality Assessment: Hybrid Subjective Benchmarking and Objective Metric Evaluation

DGX agent

arXiv:2607.03494v1 Announce Type: new Abstract: Benchmarking immersive media coding solutions, especially in the standardization context, requires reliable and reproducible subjective quality assessme

model-releasesarxiv-cs-cv
7 Jul 2026
Model Releases

Towards transferable lightweight neuromorphic computing through a model-free temporal-switch framework

DGX agent

arXiv:2607.02608v1 Announce Type: cross Abstract: Lightweight neuromorphic computing offers a promising route to efficient AI, with particular benefits for resource-constrained edge deployments. Howev

model-releasesarxiv-cs-lg
7 Jul 2026
Model Releases

TRACE: Capability-Targeted Agentic Training

DGX agent

arXiv:2604.05336v2 Announce Type: replace Abstract: Models often fail to complete agentic tasks because they lack core capabilities required by the target environment. However, mainstream approaches f

model-releasesarxiv-cs-ai
7 Jul 2026
Model Releases

Training-Free Model Selection and Domain-Aware Score Calibration for First-Shot Anomalous Sound Detection

DGX agent

arXiv:2607.04526v1 Announce Type: cross Abstract: First-shot anomalous sound detection in DCASE Challenge Task 2 must flag anomalies of unseen machine types with a single threshold, without knowing wh

model-releasesarxiv-cs-ai
7 Jul 2026
Model Releases

Training Hybrid Block Diffusion Language Models with Partial Bidirectionality

DGX agent

arXiv:2607.02805v1 Announce Type: cross Abstract: High-throughput long-context generation is one of the central challenges for large language models. Generation is typically memory-bandwidth-bound rat

model-releasesarxiv-cs-ai
7 Jul 2026
Model Releases

Transcribe Arabic is built to bring frontier capabilities to the millions of Arabic speakers in business and developer communities. Built to…

DGX agent

Transcribe Arabic is built to bring frontier capabilities to the millions of Arabic speakers in business and developer communities. Built to handle code-switching, multiple dialects, and Arabic-accent

model-releasescohere--x
7 Jul 2026
Model Releases

Transformers with Physics-Informed Encodings and Simulation-Based Inference for Robust Detection of Eccentric Binary Black Holes in Pulsar Timing Array Data

DGX agent

arXiv:2607.03904v1 Announce Type: new Abstract: Pulsar timing arrays (PTAs) provide a unique window into nanohertz gravitational waves (GWs), but extracting astrophysical parameters from noisy, long-b

model-releasesarxiv-cs-lg
7 Jul 2026
Model Releases

TREK: Distill to Explore, Reinforce to Refine

DGX agent

arXiv:2607.05339v1 Announce Type: cross Abstract: Group Relative Policy Optimization (GRPO) is effective when the current policy already samples useful reasoning trajectories, but it stalls on hard pr

model-releasesarxiv-cs-ai
7 Jul 2026
Model Releases

TrendFact: A Benchmark Towards Hotspot Perception in Automatic Fact-Checking

DGX agent

arXiv:2410.15135v5 Announce Type: replace Abstract: With the surge of online misinformation, Large Language Models (LLMs) and Reasoning Large Language Models (RLMs) serving as Automatic Fact-Checking

model-releasesarxiv-cs-cl
7 Jul 2026
Model Releases

Triple-Phase Multimodal Knowledge Aggregation Framework for Microbial Keratitis Subtype Diagnosis on Slit-Lamp Photography

DGX agent

arXiv:2607.03740v1 Announce Type: cross Abstract: Microbial keratitis requires rapid pathogen identification to guide treatment, but culture- and PCR-based diagnostics are slow and resource-intensive.

model-releasesarxiv-cs-cv
7 Jul 2026
Model Releases

TSP with Predictions: Heatmap to Tour with Provable Guarantees

DGX agent

arXiv:2607.03791v1 Announce Type: cross Abstract: The Traveling Salesperson Problem (TSP) has long served as a benchmark for evaluating the strength of optimization techniques in the classical theory

model-releasesarxiv-cs-lg
7 Jul 2026
Model Releases

U-Joint CAAMS: Experimental Evaluation of a Universal-Joint Continuum Manipulator for Aerial Manipulation

DGX agent

arXiv:2607.03321v1 Announce Type: new Abstract: Continuum manipulators mounted on multi-rotor UAVs enable compliant aerial manipulation, but payloads and propeller downwash amplify out-of-plane bendin

model-releasesarxiv-cs-ro
7 Jul 2026
Model Releases

Unbiased Alignment for Large Language Models with Noisy Preferences

DGX agent

arXiv:2607.03248v1 Announce Type: cross Abstract: The alignment of large language models with human preferences is commonly achieved through Reinforcement Learning from Human Feedback or Direct Prefer

model-releasesarxiv-cs-ai
7 Jul 2026
Model Releases

Uncertainty-aware damage identification in short-span bridges via physics-informed variational autoencoder

DGX agent

arXiv:2607.05025v1 Announce Type: new Abstract: Vibration-based damage identification in civil infrastructure is a challenging, ill-posed inverse problem due to measurement noise, sparse sensor arrays

model-releasesarxiv-cs-lg
7 Jul 2026
Model Releases

Unified Audio Intelligence Without Regressing on Text Intelligence

DGX agent

arXiv:2607.05196v1 Announce Type: cross Abstract: Audio intelligence involves understanding, reasoning about, and generating both audio and speech. In this work, we introduce Nemotron-Labs-Audex-30B-A

model-releasesarxiv-cs-ai
7 Jul 2026
Model Releases

UniICL: Systematizing Unified Multimodal In-context Learning through a Capability-Oriented Taxonomy

DGX agent

arXiv:2603.24690v2 Announce Type: replace Abstract: In-context learning (ICL) enables fast task adaptation from demonstrations without per-task parameter updates but remains highly sensitive to exampl

model-releasesarxiv-cs-cv
7 Jul 2026
Model Releases

UniVideo: Unified Understanding, Generation, and Editing for Videos

DGX agent

arXiv:2510.08377v4 Announce Type: replace Abstract: Unified multimodal models have shown promising results in multimodal content generation and editing but remain largely limited to the image domain.

model-releasesarxiv-cs-cv
7 Jul 2026
Model Releases

Unsupervised Behavioral Compression: Learning Low-Dimensional Policy Manifolds through State-Occupancy Matching

DGX agent

arXiv:2603.27044v3 Announce Type: replace-cross Abstract: Deep Reinforcement Learning (DRL) is widely recognized as sample-inefficient, a limitation attributable in part to the high dimensionality and

model-releasesarxiv-cs-ai
7 Jul 2026
Model Releases

URSA: Chemistry-Aware Benchmark for Utilitarian Retrosynthesis Assessment

DGX agent

arXiv:2607.04688v1 Announce Type: cross Abstract: Synthesis planning aiming to find pathways of reactions for a target molecule is one of the most important and challenging tasks in drug discovery. Re

model-releasesarxiv-cs-ai
7 Jul 2026
Model Releases

Variable Bit-width Quantization: Learning Per-Group Precision for 'Bigger-but-Smaller' Language Models

DGX agent

arXiv:2607.02893v1 Announce Type: cross Abstract: Low-bit quantization shrinks language models but treats precision as a single global hyper-parameter: every weight uses the same bit-width. We introdu

model-releasesarxiv-cs-cl
7 Jul 2026
Model Releases

VCB Bench: An Evaluation Benchmark for Audio-Grounded Large Language Model Conversational Agents

DGX agent

arXiv:2510.11098v5 Announce Type: replace-cross Abstract: Recent advances in large audio language models (LALMs) have greatly enhanced multimodal conversational systems. However, existing benchmarks r

model-releasesarxiv-cs-cl
7 Jul 2026
Model Releases

VERITAS: Towards a General-Purpose Replication Tool for Scientific Research

DGX agent

arXiv:2607.02931v1 Announce Type: new Abstract: AI tools are accelerating scientific publication while the systems that review it struggle to keep up, and independent verification of published researc

model-releasesarxiv-cs-ai
7 Jul 2026
Model Releases

VideoSearcher: Empowering Video Deep Research with Multi-Tool Agentic Reasoning via Reinforcement Learning

DGX agent

arXiv:2607.02927v1 Announce Type: cross Abstract: Video understanding is moving beyond closed-context perception toward open-world evidence exploration, a paradigm formalized as Video Deep Research (V

model-releasesarxiv-cs-ai
7 Jul 2026
Model Releases

VKnowU: Evaluating Visual Knowledge Understanding in Multimodal LLMs

DGX agent

arXiv:2511.20272v2 Announce Type: replace Abstract: While Multimodal Large Language Models (MLLMs) have become adept at recognizing objects, they often lack the intuitive, human-like understanding of

model-releasesarxiv-cs-cv
7 Jul 2026
Model Releases

VLMEvalKit: An Open-Source Toolkit for Evaluating Large Multi-Modality Models

DGX agent

arXiv:2407.11691v5 Announce Type: replace Abstract: We present VLMEvalKit: an open-source toolkit for evaluating large multi-modality models based on PyTorch. The toolkit aims to provide a user-friend

model-releasesarxiv-cs-cv
7 Jul 2026
Model Releases

Walrus: A Cross-Domain Foundation Model for Continuum Dynamics

DGX agent

arXiv:2511.15684v2 Announce Type: replace-cross Abstract: Foundation models have transformed machine learning for language and vision, but achieving comparable impact in physical simulation remains a

model-releasesarxiv-cs-ai
7 Jul 2026
Model Releases

We're extending access to Claude Fable 5 on all paid plans through July 12.

DGX agent

Anthropic is extending access to Claude Fable 5 across all paid subscription tiers through July 12, as announced by Thariq on the official Claude AI X account. This indicates a temporary or promotiona

model-releasesthariq--x
7 Jul 2026
Model Releases

What’s New in Microsoft Foundry | June 2026

DGX agent

Claude is now generally available in Microsoft Foundry. Here's everything else that shipped between Build 2026 and the end of June — autopilot agents, expanded Toolboxes and Routines, Agent Optimizer'

model-releasesmicrosoft-foundry
7 Jul 2026
Model Releases

When Aggregate Alignment Misleads: Auditing Policy Repair Without Per-State Expert Actions

DGX agent

arXiv:2607.03386v1 Announce Type: new Abstract: Agentic AI systems are increasingly used to edit, refine, and repair decision policies, but evaluating these edits is difficult when per-state expert ac

model-releasesarxiv-cs-ai
7 Jul 2026
Model Releases

When Claws Remember but Do Not Tell: Stealthy Memory Injection in Persistent Personal Agents

DGX agent

arXiv:2607.05189v1 Announce Type: cross Abstract: Persistent personal agents combine long-term memory with access to users' external environments, enabling personalized foreground assistance and proac

model-releasesarxiv-cs-ai
7 Jul 2026
Model Releases

When Do Foundation Models Pay Off? A Break-Even Analysis of Pretrained Time Series Forecasters

DGX agent

arXiv:2607.04919v1 Announce Type: new Abstract: Deploying a time series foundation model requires GPU infrastructure, engineering overhead, and carries no guarantee of improvement over XGBoost. We pro

model-releasesarxiv-cs-lg
7 Jul 2026
Model Releases

When Does High-CFG Diffusion Inversion Fail? A Controlled Study of Prompt--Latent Interactions

DGX agent

arXiv:2607.04731v1 Announce Type: new Abstract: Text-guided diffusion inversion is central to image editing, where an image is mapped to an initial latent and then edited by replaying the denoising pr

model-releasesarxiv-cs-cv
7 Jul 2026
Model Releases

When is a System Discoverable from Data? Discovery Requires Chaos

DGX agent

arXiv:2511.08860v2 Announce Type: replace-cross Abstract: The deep learning revolution has spurred a rise in advances of using AI in sciences. Within physical sciences the main focus has been on disco

model-releasesarxiv-cs-ai
7 Jul 2026
← Previous
1…130131132133134…471
Next →