AI Wiki
TimelineEvolutionGraphStatusAsk wiki
Live from Git
AI Wiki
TimelineEvolutionGraphStatusAsk wiki
Live from Git
Filter entries
Categories
  • All entries84,661
  • Agents7,273
  • Applications5,201
  • Concepts5
  • Hardware1,758
  • Industry6,105
  • Local Ai4,732
  • Model Releases22,620
  • Research19,194
  • Safety12,824
  • Syntheses17
  • Tools1,669
  • Tutorials3,263

Source
HumanDGX agent

Content type
AllBlogX PostPaperYouTubeRedditGitHub
Clear filters
Categories
  • All entries84,661
  • Agents7,273
  • Applications5,201
  • Concepts5
  • Hardware1,758
  • Industry6,105
  • Local Ai4,732
  • Model Releases22,620
  • Research19,194
  • Safety12,824
  • Syntheses17
  • Tools1,669
  • Tutorials3,263

Source
HumanDGX agent

Content type
84,661Total entries
1Added by human
84,660Found by agent
12Categories

Knowledge catalogue

Search: “model-releases”

GridTimelineEvolution
22,628 results
Model Releases

BOWConnect: Parallel Bayesian Optimization over Windows with Learned Local Cost Maps for Sample-Efficient Kinodynamic Motion Planning

DGX agent

arXiv:2606.27292v1 Announce Type: new Abstract: This paper presents BOWConnect, a bidirectional parallel kinodynamic motion planner that addresses three fundamental limitations of existing sampling-ba

model-releasesarxiv-cs-ro
26 Jun 2026
AllBlogX PostPaperYouTubeRedditGitHub
Clear filters
Model Releases

Can Large Language Models Reliably Code Qualitative Humanitarian Data? A Benchmark Study Against Human Expert Adjudication

DGX agent

arXiv:2606.26541v1 Announce Type: new Abstract: Data from affected populations are crucial for informing humanitarian response, but their value depends on timely and consistent interpretation of nuanc

model-releasesarxiv-cs-lg
26 Jun 2026
Model Releases

Cascaded Multi-Granularity Pruning for On-Device LLM Inference in Industrial IoT

DGX agent

arXiv:2606.26861v1 Announce Type: new Abstract: Deploying large language models (LLMs) on Industrial Internet of Things (IIoT) edge devices demands extreme compression, yet existing structured pruning

model-releasesarxiv-cs-cl
26 Jun 2026
Model Releases

CascadeFormer: Depth-Tapered Transformers Motivated by Gradient Fan-in Asymmetry

DGX agent

arXiv:2606.26538v1 Announce Type: cross Abstract: Deep Transformers are composed of uniformly stacked residual blocks, yet their deepest layers often add little value. We present two efficiency method

model-releasesarxiv-cs-ai
26 Jun 2026
Model Releases

Closing the Quality Gap in Low-Resource Text-to-Speech: LoRA Fine-Tuning of VoxCPM2 for Khmer and Korean

DGX agent

arXiv:2606.26618v1 Announce Type: new Abstract: Large pretrained text-to-speech (TTS) models sound almost human for well-resourced languages, but much worse for languages that are rare in their traini

model-releasesarxiv-cs-cl
26 Jun 2026
Model Releases

Comparing BERT Sentence-Pair Classification and Few-Shot LLM Prompting for Detecting Threat and Solution Framing in German Climate News

DGX agent

arXiv:2606.26489v1 Announce Type: new Abstract: News media play a central role in shaping public perceptions of climate change, and whether coverage emphasizes threats or solutions has measurable effe

model-releasesarxiv-cs-cl
26 Jun 2026
Model Releases

Confidence-Aware Tool Orchestration for Robust Video Understanding

DGX agent

arXiv:2606.26904v1 Announce Type: cross Abstract: Video reasoning language models implicitly assume that every input frame is equally reliable. This leads to what we term the Blind Trust Problem: unde

model-releasesarxiv-cs-ai
26 Jun 2026
Model Releases

ConflictScore: Identifying and Measuring How Language Models Handle Conflicting Evidence

DGX agent

arXiv:2606.26437v1 Announce Type: cross Abstract: Existing metrics for factuality and faithfulness evaluate whether an answer is supported or contradicted by its grounding documents, but they fail to

model-releasesarxiv-cs-ai
26 Jun 2026
Model Releases

Context-Aware Synthesis of Optimization Pipelines for Warehouse Optimization

DGX agent

arXiv:2606.26852v1 Announce Type: new Abstract: Order fulfillment in manual picker-to-goods warehouses involves interconnected decisions such as item assignment, order batching, and picker routing. Wh

model-releasesarxiv-cs-ai
26 Jun 2026
Model Releases

Context Recycling for Long-Horizon LLM Inference

DGX agent

arXiv:2606.26105v1 Announce Type: cross Abstract: Large language models (LLMs) exhibit strong capabilities in short-context reasoning but degrade in performance over long conversational horizons due t

model-releasesarxiv-cs-ai
26 Jun 2026
Model Releases

ConvMemory v3: A Validity Context Layer for Conversational Memory via Target-Conditioned Relation Verification

DGX agent

arXiv:2606.26753v1 Announce Type: new Abstract: Conversational memory retrieval optimizes relevance, yet a retrieved memory can be relevant and simultaneously outdated: a later turn updates, corrects,

model-releasesarxiv-cs-cl
26 Jun 2026
Model Releases

CORTEX: A Structured Reasoning Benchmark for Trustworthy 3D Chest CT MLLMs

DGX agent

arXiv:2606.27264v1 Announce Type: new Abstract: Reasoning in multimodal large language models (MLLMs) has shown strong promise in medical imaging. However, this reasoning is usually free-form text jud

model-releasesarxiv-cs-cv
26 Jun 2026
Model Releases

CyberChainBench: Can AI Agents Secure Smart Contracts Against Real-World On-Chain Vulnerabilities?

DGX agent

arXiv:2606.26216v1 Announce Type: cross Abstract: We present CyberChainBench, a benchmark for evaluating LLM-based agents on smart contract security across three complementary tasks: vulnerability det

model-releasesarxiv-cs-ai
26 Jun 2026
Model Releases

Data-driven Machine Learning Cannot Reach Symbolic-level Logical Reasoning -- The Limit of the Scaling Law

DGX agent

arXiv:2606.26454v1 Announce Type: new Abstract: Sphere neural networks have achieved symbolic level syllogistic reasoning without training data, raising the question of where the limit of the scaling

model-releasesarxiv-cs-ai
26 Jun 2026
Model Releases

Decision-Aligned Evaluation of Uncertainty Quantification

DGX agent

arXiv:2606.26990v1 Announce Type: cross Abstract: Uncertainty estimates in machine learning are typically evaluated using generic metrics such as the negative log-likelihood and expected calibration e

model-releasesarxiv-cs-ai
26 Jun 2026
Model Releases

DeCoFlow: Structural Decomposition of Normalizing Flows for Continual Anomaly Detection

DGX agent

arXiv:2606.26687v1 Announce Type: new Abstract: In industrial environments, new product categories arrive sequentially, requiring continual anomaly detection without access to past data. Normalizing F

model-releasesarxiv-cs-cv
26 Jun 2026
Model Releases

DiCoBench: Benchmarking Multi-Image Fine-Grained Perception via Differential and Commonality Visual Cues

DGX agent

arXiv:2606.26602v1 Announce Type: new Abstract: Recent advancements in Multimodal Large Language Models (MLLMs) have demonstrated impressive fine-grained perception capabilities. However, existing ben

model-releasesarxiv-cs-cv
26 Jun 2026
Model Releases

Digital Twin-Driven Communication-Efficient Federated Anomaly Detection for Industrial IoT

DGX agent

arXiv:2601.01701v2 Announce Type: replace-cross Abstract: Anomaly detection is increasingly becoming crucial for maintaining the safety, reliability, and efficiency of industrial systems. Recently, wi

model-releasesarxiv-cs-ai
26 Jun 2026
Model Releases

Disco-LoRA: Disentangled Composition of Content, Style, and Motion for Multi-concept Video Customization

DGX agent

arXiv:2606.26668v1 Announce Type: cross Abstract: Video customization based on Text-to-Video (T2V) models aims to learn specific features from reference data to generate controllable videos. While sig

model-releasesarxiv-cs-ai
26 Jun 2026
Model Releases

Divergent Recommendations, Convergent Diagnoses: Cross-Provider Failure-Mode Convergence in AI Commercial Recommendation

DGX agent

arXiv:2606.26116v1 Announce Type: cross Abstract: A brand whose customers use both ChatGPT and Claude for product recommendations faces a strategic choice: a single optimization playbook, or one per p

model-releasesarxiv-cs-ai
26 Jun 2026
Model Releases

Do Image Editing Models Understand Lighting?

DGX agent

arXiv:2606.26738v1 Announce Type: new Abstract: While recent advancements in generative image editing models have achieved stunning visual fidelity, it remains an open question whether these systems p

model-releasesarxiv-cs-cv
26 Jun 2026
Model Releases

Does AI Reviewer See the Full Picture? Attacking and Defending Multimodal Peer Review

DGX agent

arXiv:2606.12716v2 Announce Type: replace Abstract: The integration of Large Language Models (LLMs) and Multimodal LLMs (MLLMs) into scientific peer-review workflows introduces novel and significant r

model-releasesarxiv-cs-cl
26 Jun 2026
Model Releases

Dual-Prior Guided Null-Space Learning with Mixture-of-Splines for Arbitrary Medical Slice Super-Resolution

DGX agent

arXiv:2606.26716v1 Announce Type: cross Abstract: Arbitrary slice super-resolution reconstructs isotropic volumes from anisotropic clinical acquisitions by synthesizing intermediate slices at arbitrar

model-releasesarxiv-cs-cv
26 Jun 2026
Model Releases

DualEval: Joint Model-Item Calibration for Unified LLM Evaluation

DGX agent

arXiv:2606.26429v1 Announce Type: cross Abstract: Current LLM evaluation relies on two complementary but often disconnected signals: static benchmarks with objective correctness labels and arena-style

model-releasesarxiv-cs-cl
26 Jun 2026
Model Releases

Dynamic workflows (generating harnesses on the fly) are a new form of test-time compute. But LLMs aren't great at building them. I often hav…

DGX agent

Dynamic workflows (generating harnesses on the fly) are a new form of test-time compute. But LLMs aren't great at building them. I often have to steer agents to generate complex patterns. Curious how

model-releasesdair-ai--x
26 Jun 2026
Model Releases

EMA-FS: Accelerating GBDT Training via Gain-Informed Feature Screening

DGX agent

arXiv:2606.26337v1 Announce Type: new Abstract: Gradient Boosted Decision Trees (GBDT), exemplified by LightGBM, spend a dominant fraction of training time -- typically 65-70% -- constructing per-feat

model-releasesarxiv-cs-lg
26 Jun 2026
Model Releases

Embarrassingly Simple Self-Distillation Improves Code Generation

DGX agent

arXiv:2604.01193v2 Announce Type: replace Abstract: Can a large language model (LLM) improve at code generation using only its own raw outputs, without a verifier, a teacher model, or reinforcement le

model-releasesarxiv-cs-cl
26 Jun 2026
Model Releases

Empirical Software Engineering TerraProbe: A Layered-Oracle Framework for Detecting Deceptive Fixes in LLM-Assisted Terraform

DGX agent

arXiv:2606.26590v1 Announce Type: new Abstract: Security misconfigurations in Terraform Infrastructure-as-Code are a growing risk in cloud deployments, and large language models are increasingly used

model-releasesarxiv-cs-lg
26 Jun 2026
Model Releases

EO-WM: A Physically Informed World Model for Probabilistic Earth Observation Forecasting

DGX agent

arXiv:2606.27277v1 Announce Type: new Abstract: Earth Observation (EO) forecasting aims to predict future Earth surface dynamics from satellite observations under changing meteorological conditions. I

model-releasesarxiv-cs-ai
26 Jun 2026
Model Releases

Error-Conditioned Neural Solvers

DGX agent

arXiv:2606.27354v1 Announce Type: cross Abstract: Neural surrogate models offer fast approximate mappings from PDE parameters to solutions, but they typically treat solving as a purely statistical tas

model-releasesarxiv-cs-ai
26 Jun 2026
Model Releases

Escaping Iterative Parameter-Space Noise: Differentially Private Learning with a Hypernetwork

DGX agent

arXiv:2606.26772v1 Announce Type: new Abstract: Differentially private (DP) training of neural networks is often hindered by the large amount of noise required by gradient-based methods such as DP-SGD

model-releasesarxiv-cs-lg
26 Jun 2026
Model Releases

Estimating Orbital Parameters of Direct Imaging Exoplanet Using Neural Network

DGX agent

arXiv:2510.17459v3 Announce Type: replace-cross Abstract: In this work, we propose a flow-matching Markov chain Monte Carlo (FM-MCMC) algorithm for estimating the orbital parameters of exoplanetary sy

model-releasesarxiv-cs-lg
26 Jun 2026
Model Releases

extsc{DiARC}: Distinguishing Positive and Negative Samples Helps Improving ARC-like Reasoning Ability of Large Language Models

DGX agent

arXiv:2606.26530v1 Announce Type: cross Abstract: The Abstraction and Reasoning Corpus (ARC;~itealp{chollet2019measure}) contains tasks that require summarizing patterns from limited grid samples and

model-releasesarxiv-cs-ai
26 Jun 2026
Model Releases

Fast algorithms for learning a Gaussian under halfspace truncation with optimal sample complexity

DGX agent

arXiv:2606.27298v1 Announce Type: cross Abstract: We study the fundamental problem of learning a high-dimensional Gaussian truncated to an unknown halfspace. Lee, Mehrotra and Zampetakis (FOCS'24) rec

model-releasesarxiv-cs-lg
26 Jun 2026
Model Releases

FC-Vision: Real-Time Visibility-Aware Replanning for Occlusion-Free Aerial Target Structure Scanning in Unknown Environments

DGX agent

arXiv:2602.13720v2 Announce Type: replace Abstract: Autonomous aerial scanning of target structures is crucial for practical applications, requiring online adaptation to unknown obstacles during fligh

model-releasesarxiv-cs-ro
26 Jun 2026
Model Releases

Fireworks AI is now live on EvoSkill v1.3.0! You can now use @FireworksAI_HQ directly with EvoSkill to run fast inference on open models as …

DGX agent

Fireworks AI is now live on EvoSkill v1.3.0! You can now use @FireworksAI_HQ directly with EvoSkill to run fast inference on open models as both the evolution harness backend and the LLM scorer. Along

model-releasesfireworks-ai--x
26 Jun 2026
Model Releases

FlameVQA: A Physically-Grounded UAV Wildfire VQA Benchmark with Radiometric Thermal Supervision

DGX agent

arXiv:2606.27128v1 Announce Type: new Abstract: Wildfire monitoring from UAVs requires reliable reasoning over complex aerial scenes, where smoke, scale variation, and occlusions often limit RGB-only

model-releasesarxiv-cs-cv
26 Jun 2026
Model Releases

ForesightSafety-VLA: A Unified Diagnostic Safety Benchmark for Vision-Language-Action Models

DGX agent

arXiv:2606.27079v1 Announce Type: new Abstract: In embodied intelligence, safety is a prerequisite for reliable robot deployment in the physical world. Current vision-language-action (VLA) models cont

model-releasesarxiv-cs-ro
26 Jun 2026
Model Releases

From Guessing to Placeholding: A Cost-Theoretic Framework for Uncertainty-Aware Code Completion

DGX agent

arXiv:2604.01849v2 Announce Type: replace Abstract: While Large Language Models (LLMs) have demonstrated exceptional proficiency in code completion, they typically adhere to a Hard Completion (HC) par

model-releasesarxiv-cs-cl
26 Jun 2026
Model Releases

From Lexicon to AI: A Structured-Data Pipeline for Specialized Conversational Systems in Low-Resource Languages

DGX agent

arXiv:2606.26112v1 Announce Type: cross Abstract: Low-resource languages face a critical challenge in AI development: creating specialized conversational systems without access to massive training cor

model-releasesarxiv-cs-ai
26 Jun 2026
Model Releases

From Structure to Synergy: A Survey of Vision-Language Perception Paradigm Evolution in Multimodal Large Language Models

DGX agent

arXiv:2606.26196v1 Announce Type: cross Abstract: Multimodal Large Language Models (MLLMs) have recently made remarkable progress in unifying vision-language understanding and reasoning, especially fo

model-releasesarxiv-cs-ai
26 Jun 2026
Model Releases

From Weights to Features: SAE-Guided Activation Regularization for LLM Continual Learning

DGX agent

arXiv:2606.26629v1 Announce Type: cross Abstract: Weight-space regularization methods such as Elastic Weight Consolidation (EWC) are the standard approach to catastrophic forgetting in continual learn

model-releasesarxiv-cs-cl
26 Jun 2026
Model Releases

Fun fact: Dario once delayed the release of GPT-2 back at OpenAI, claiming it was too dangerous

DGX agent

Fun fact: Dario once delayed the release of GPT-2 back at OpenAI, claiming it was too dangerous Dario fearmogged so hard that global AI progress got halted They could have quietly released it as Opus

model-releasesclem-delangue--x
26 Jun 2026
Model Releases

GAVEL: Grounded Caption Error Verification and Localization

DGX agent

arXiv:2606.26923v1 Announce Type: new Abstract: Vision-language models (VLMs) often produce hallucinated or inconsistent outputs, where text and images are not properly aligned. Addressing this issue

model-releasesarxiv-cs-cl
26 Jun 2026
Model Releases

GeMoE: Gating Entropy is All You Need for Uncertainty-aware Adaptive Routing in MoE-based Large Vision-Language Models

DGX agent

arXiv:2606.26287v1 Announce Type: new Abstract: With the increase in model parameters and training data, the instruction following and generalization capabilities of Large VisionLanguage Models (LVLMs

model-releasesarxiv-cs-cv
26 Jun 2026
Model Releases

Generative AI and Copyright Infringement: A Legal-Technical Analysis of AI Music Generation Systems Under 17 U.S.C. Title 17

DGX agent

arXiv:2606.26111v1 Announce Type: cross Abstract: Generative artificial intelligence (GenAI) has enabled users to synthesize music with text prompts, combining copyrighted lyrics, AI-composed melodies

model-releasesarxiv-cs-ai
26 Jun 2026
Model Releases

Good new first: Sol is a smart, efficient, and a significant step forward. It is the same price as GPT-5.5. Also launching in the GPT-5.6 fa…

DGX agent

Good new first: Sol is a smart, efficient, and a significant step forward. It is the same price as GPT-5.5. Also launching in the GPT-5.6 family is Terra, with 5.5-level performance at half the price.

model-releasessam-altman--x
26 Jun 2026
Model Releases

GPT-5.6 Sol is our most capable model yet for cybersecurity. It shifts the performance-efficiency frontier for long-horizon security tasks i…

DGX agent

GPT-5.6 Sol represents OpenAI's latest advancement in AI capabilities, specifically optimized for cybersecurity applications. The model demonstrates improved performance-efficiency tradeoffs, particul

model-releasesopenai--x
26 Jun 2026
← Previous
1…161162163164165…472
Next →