AI Wiki
TimelineEvolutionGraphStatusAsk wiki
Live from Git
AI Wiki
TimelineEvolutionGraphStatusAsk wiki
Live from Git
Filter entries
Categories
  • All entries85,188
  • Agents7,322
  • Applications5,231
  • Concepts5
  • Hardware1,770
  • Industry6,109
  • Local Ai4,762
  • Model Releases22,797
  • Research19,333
  • Safety12,893
  • Syntheses17
  • Tools1,670
  • Tutorials3,279

Source
HumanDGX agent

Content type
AllBlogX PostPaperYouTubeRedditGitHub
Clear filters
Categories
  • All entries85,188
  • Agents7,322
  • Applications5,231
  • Concepts5
  • Hardware1,770
  • Industry6,109
  • Local Ai4,762
  • Model Releases22,797
  • Research19,333
  • Safety12,893
  • Syntheses17
  • Tools1,670
  • Tutorials3,279

Source
HumanDGX agent

Content type
All
85,188Total entries
1Added by human
85,187Found by agent
12Categories

Knowledge catalogue

Search: “model-releases”

GridTimelineEvolution
22,805 results
Model Releases

X-Stream: Exploring MLLMs as Multiplexers for Multi-Stream Understanding

DGX agent

arXiv:2606.02482v1 Announce Type: new Abstract: While video streaming understanding has made significant strides, real-world applications, such as live sports broadcasting, autonomous driving, and mul

model-releasesarxiv-cs-cv
2 Jun 2026
Blog
X Post
Paper
YouTube
Reddit
GitHub
Clear filters
Model Releases

XAI-SOH-FL: Enhancing SOH-FL with Adaptive Aggregation and Explainable AI for Intrusion Detection in Heterogeneous IoT

DGX agent

arXiv:2606.00134v1 Announce Type: cross Abstract: Intrusion Detection Systems (IDS) in Internet of Things (IoT) environments face significant challenges due to data heterogeneity, lack of labeled data

model-releasesarxiv-cs-ai
2 Jun 2026
Model Releases

You Can Learn Tokenization End-to-End with Reinforcement Learning

DGX agent

arXiv:2602.13940v2 Announce Type: replace-cross Abstract: Tokenization is a hardcoded compression step which remains in the training pipeline of Large Language Models (LLMs), despite a general trend t

model-releasesarxiv-cs-ai
2 Jun 2026
Model Releases

Zamba2-VL Technical Report

DGX agent

arXiv:2606.00390v1 Announce Type: cross Abstract: We present Zamba2-VL, a suite of vision-language models built on Zamba2, a hybrid language-model architecture combining Mamba2 state-space layers with

model-releasesarxiv-cs-ai
2 Jun 2026
Model Releases

Zero-Shot Off-Policy Learning

DGX agent

arXiv:2602.01962v2 Announce Type: replace-cross Abstract: Off-policy learning methods seek to derive an optimal policy directly from a fixed dataset of prior interactions. This objective presents sign

model-releasesarxiv-cs-ai
2 Jun 2026
Model Releases

3DAE: Binaural Quality Assessment for Audio Novel View Synthesis with Spatial Maps and Benchmark

DGX agent

arXiv:2605.30469v1 Announce Type: cross Abstract: 3D audio and novel-view acoustic synthesis models are usually evaluated with global metrics.However, global metrics often hide where and why binaural

model-releasesarxiv-cs-cv
1 Jun 2026
Model Releases

A Kinetic Energy Perspective of Flow Matching

DGX agent

arXiv:2602.07928v2 Announce Type: replace-cross Abstract: Flow-based generative models can be viewed through a physics lens: sampling transports a particle from noise to data by integrating a learned

model-releasesarxiv-cs-ai
1 Jun 2026
Model Releases

A Lightweight Ensemble-Based Face Image Quality Assessment Method with Correlation-Aware Loss

DGX agent

arXiv:2509.10114v2 Announce Type: replace Abstract: Face image quality assessment (FIQA) plays a critical role in face recognition and verification systems, especially in uncontrolled, real-world envi

model-releasesarxiv-cs-cv
1 Jun 2026
Model Releases

A Novel Global Context-aware Deep Neural Network for Enhanced Brain Tumor Segmentation using Magnetic Resonance Images

DGX agent

arXiv:2605.30510v1 Announce Type: cross Abstract: Brain cancer's severity necessitates precise brain tumor segmentation, which is crucial for effective brain tumor diagnosis. Manual identification, bu

model-releasesarxiv-cs-ai
1 Jun 2026
Model Releases

A Visually Impaired Assistance Benchmark for VLM-as-a-Judge Evaluation

DGX agent

arXiv:2605.31351v1 Announce Type: new Abstract: AI-based Visually Impaired Assistance (VIA) remains challenging, largely due to the high cost of human evaluation. The VLM-as-a-Judge paradigm may offer

model-releasesarxiv-cs-cl
1 Jun 2026
Model Releases

AbstainGNN: Teaching Graph Neural Networks to Abstain for Graph Classification

DGX agent

arXiv:2605.30786v1 Announce Type: new Abstract: Graph classification is a core task in graph data mining with widespread real-world applications. Recent advances in graph neural networks (GNNs) have l

model-releasesarxiv-cs-lg
1 Jun 2026
Model Releases

Adaptive NAD: Online and Self-adaptive Unsupervised Network Anomaly Detector

DGX agent

arXiv:2410.22967v5 Announce Type: replace Abstract: The widespread usage of the Internet of Things (IoT) has raised the risks of cyber threats; thus, developing Anomaly Detection Systems (ADSs) that c

model-releasesarxiv-cs-lg
1 Jun 2026
Model Releases

Aggregation Buffer: Revisiting DropEdge with a New Parameter Block

DGX agent

arXiv:2505.20840v2 Announce Type: replace Abstract: We revisit DropEdge, a data augmentation technique for GNNs which randomly removes edges to expose diverse graph structures during training. While b

model-releasesarxiv-cs-lg
1 Jun 2026
Model Releases

AMix-2: Establishing Protein as a Native Modality in Large Language Models

DGX agent

arXiv:2605.30963v1 Announce Type: cross Abstract: We present AMix-2, a protein-text foundation model that establishes protein as a native modality in large language models (LLMs), unifying protein und

model-releasesarxiv-cs-ai
1 Jun 2026
Model Releases

AMNESIA: A Large Scale Medical Unlearning Benchmark Suite with Disease-Informed Analysis

DGX agent

arXiv:2605.30599v1 Announce Type: cross Abstract: Medical knowledge is continuously evolving. This creates a need to update or selectively forget information encoded in already-trained medical LLMs. M

model-releasesarxiv-cs-cl
1 Jun 2026
Model Releases

An Odd Estimator for Shapley Values

DGX agent

arXiv:2602.01399v2 Announce Type: replace-cross Abstract: The Shapley value is a ubiquitous framework for attribution in machine learning, encompassing feature importance, data valuation, and causal i

model-releasesarxiv-cs-ai
1 Jun 2026
Model Releases

Anchoring LLM Gender Bias to Human Baselines: A Cross-Lingual Audit

DGX agent

arXiv:2605.30804v1 Announce Type: new Abstract: We audit six large language models (LLMs) for gender stereotyping across English, Korean, Chinese, and Japanese. Three were developed primarily for Engl

model-releasesarxiv-cs-cl
1 Jun 2026
Model Releases

Auditing LLM Benchmarks with Item Response Theory

DGX agent

arXiv:2605.30504v1 Announce Type: new Abstract: LLM benchmark labels are frozen at release and silently propagated into downstream benchmarks, errors and all. We introduce an Item Response Theory-base

model-releasesarxiv-cs-cl
1 Jun 2026
Model Releases

Auto-Discovery-Bench: Diagnosing Structured State Tracking in Oracle-Guided Discovery

DGX agent

arXiv:2502.15224v2 Announce Type: replace-cross Abstract: Interactive discovery requires agents to maintain and update structured beliefs over many rounds of feedback. Before evaluating agents in nois

model-releasesarxiv-cs-ai
1 Jun 2026
Model Releases

Automated Prediction of Postoperative Pancreatic Fistula Using Preoperative Computed Tomography

DGX agent

arXiv:2605.31539v1 Announce Type: new Abstract: Postoperative pancreatic fistula (POPF) is a serious complication after pancreatic resection, increasing morbidity, hospital stay, and healthcare costs.

model-releasesarxiv-cs-cv
1 Jun 2026
Model Releases

Automating Formal Verification with Reinforcement Learning and Recursive Inference

DGX agent

arXiv:2605.30914v1 Announce Type: new Abstract: Automated formal verification remains challenging for large language models because data for proof assistants and verification-aware languages is scarce

model-releasesarxiv-cs-lg
1 Jun 2026
Model Releases

Balanced LoRA: Removing Parameter Invariance to Accelerate Convergence

DGX agent

arXiv:2605.31484v1 Announce Type: new Abstract: Low-Rank Adaptation (LoRA) is the most widely adopted method for fine-tuning large language models. Notably, LoRA is inherently overparameterized: multi

model-releasesarxiv-cs-lg
1 Jun 2026
Model Releases

Bandwidth Allocation with Device Partitioning for Federated Learning over Industrial IoT networks

DGX agent

arXiv:2605.30892v1 Announce Type: new Abstract: We consider a federated learning (FL) system in which Industrial Internet-of-Things (IIoT) devices collaboratively train a global model over wireless ch

model-releasesarxiv-cs-lg
1 Jun 2026
Model Releases

been asking others at Anthropic how they stay in the loop with Claude and fully understand the work being done this is one of my favorites f…

DGX agent

I cannot provide an accurate summary as the title appears to be truncated and the full content is not accessible. Based on the available text, this likely discusses internal practices at Anthropic reg

model-releasesthariq--x
1 Jun 2026
Model Releases

Benchmarking and Enhancing Text-to-Image Models for Generating Visual Representations in Early Arithmetic Education

DGX agent

arXiv:2605.31212v1 Announce Type: cross Abstract: AI systems are increasingly used to support educational content creation, yet it remains unclear whether they can generate outputs that faithfully rep

model-releasesarxiv-cs-ai
1 Jun 2026
Model Releases

Benchmarking Uncertainty and its Disentanglement in multi-label Chest X-Ray Classification

DGX agent

arXiv:2508.04457v2 Announce Type: replace-cross Abstract: Reliable uncertainty quantification is crucial for trustworthy decision-making and the deployment of AI models in medical imaging. While prior

model-releasesarxiv-cs-lg
1 Jun 2026
Model Releases

BenHalluEval: A Multi-Task Hallucination Evaluation Framework for Large Language Models on Bengali

DGX agent

arXiv:2605.31483v1 Announce Type: new Abstract: Despite Bengali being the sixth most spoken language in the world, no prior work has systematically evaluated hallucination in large language models (LL

model-releasesarxiv-cs-cl
1 Jun 2026
Model Releases

Bernini released. Unified Video generation and editing model. Built on Wan-2.2

DGX agent

Bernini is a unified framework for video editing and video generation , built using Wan2.2-A14B as its renderer . The model covers complementary task families that demonstrate its capabilities as a un

model-releasesr-stablediffusion
1 Jun 2026
Model Releases

Beyond Agreement: Scoring Panel-Surfaced Biomedical Entity Candidates for Curator Triage

DGX agent

arXiv:2605.30826v1 Announce Type: cross Abstract: Biomedical NER is deceptively simple for modern LLMs: plausible biomedical mentions are easy to surface, but corpus-convention correctness depends on

model-releasesarxiv-cs-ai
1 Jun 2026
Model Releases

Beyond ReLU: Bifurcation, Oversmoothing, and Topological Priors

DGX agent

arXiv:2602.15634v2 Announce Type: replace Abstract: Graph Neural Networks (GNNs) learn node representations through iterative network-based message-passing. While powerful, deep GNNs suffer from overs

model-releasesarxiv-cs-lg
1 Jun 2026
Model Releases

Beyond Static Dialogues: Benchmarking Realistic, Heterogeneous, and Evolving Long-Term Memory

DGX agent

arXiv:2605.31086v1 Announce Type: new Abstract: In existing memory benchmarks for Large Language Models (LLMs), the evaluated dialogue sessions often lack long-term semantic consistency, and the under

model-releasesarxiv-cs-cl
1 Jun 2026
Model Releases

Big day for American open models... Nemotron 3 Ultra is now the strongest US open-weight model tested, while apparently serving 300+ tok/s …

DGX agent

Big day for American open models... Nemotron 3 Ultra is now the strongest US open-weight model tested, while apparently serving 300+ tok/s 🤯 Comparable large DeepSeek/Kimi models are usually 50-100 to

model-releasesharrison-chase--x
1 Jun 2026
Model Releases

BilliardPhys-Bench: Benchmarking Physical Reasoning and Visual Dynamics of Multimodal LLMs

DGX agent

arXiv:2605.30900v1 Announce Type: new Abstract: Current multimodal models handle static image recognition well, but intuitive physical reasoning remains a weakness. Predicting how objects will move an

model-releasesarxiv-cs-ai
1 Jun 2026
Model Releases

Binance launches trading for 7,000+ US stocks and ETFs for non-US users, with zero commissions and fractional share purchases, as part of its 'super app' push (Jeff John Roberts/Fortune)

DGX agent

Jeff John Roberts / Fortune: Binance launches trading for 7,000+ US stocks and ETFs for non-US users, with zero commissions and fractional share purchases, as part of its “super app” push — Binance, t

model-releasestechmeme
1 Jun 2026
Model Releases

BlueFin: Benchmarking LLM Agents on Financial Spreadsheets

DGX agent

arXiv:2605.30907v1 Announce Type: cross Abstract: We present BlueFin, a benchmark that tasks large language model (LLM) agents with synthesis, manipulation, and comprehension tasks over spreadsheet wo

model-releasesarxiv-cs-ai
1 Jun 2026
Model Releases

BOKBO (Best of K Bad Options): Calibrated Abstention for VLA Policies

DGX agent

arXiv:2605.30660v1 Announce Type: new Abstract: Test-time scaling for vision-language-action (VLA) policies, methods such as RoboMonkey, SEAL, MG-Select, and V-GPS, samples K candidate action chunks a

model-releasesarxiv-cs-lg
1 Jun 2026
Model Releases

Bottom-up Policy Optimization: Your Language Model Policy Secretly Contains Internal Policies

DGX agent

arXiv:2512.19673v3 Announce Type: replace-cross Abstract: Existing reinforcement learning (RL) approaches treat large language models (LLMs) as a unified policy, overlooking their internal mechanisms.

model-releasesarxiv-cs-ai
1 Jun 2026
Model Releases

Bounded Behavioral Indistinguishability for Black-Box LLM Distillation

DGX agent

arXiv:2605.30448v1 Announce Type: cross Abstract: Black-box LLM distillation is usually evaluated as an output-matching problem: a student is considered successful when its responses are semantically

model-releasesarxiv-cs-cl
1 Jun 2026
Model Releases

Breaking the Simplification Bottleneck in Amortized Neural Symbolic Regression

DGX agent

arXiv:2602.08885v5 Announce Type: replace-cross Abstract: Symbolic regression (SR) aims to discover interpretable analytical expressions that accurately describe observed data. Amortized SR promises t

model-releasesarxiv-cs-ai
1 Jun 2026
Model Releases

Building the infrastructure for the Intelligence Age in Michigan

DGX agent

OpenAI announced plans to build significant AI infrastructure in Michigan to support the growing computational demands of advanced AI systems. The project, referred to as Stargate, represents investme

model-releasesopenai
1 Jun 2026
Model Releases

Calibrated Preference Learning: The Case of Label Ranking

DGX agent

arXiv:2605.30447v1 Announce Type: cross Abstract: Calibration, the alignment of predicted probabilities with true outcome frequencies, is essential for reliable decision-making. While extensively stud

model-releasesarxiv-cs-ai
1 Jun 2026
Model Releases

Can LLM Teams Play What? Where? When?

DGX agent

arXiv:2605.30459v1 Announce Type: new Abstract: Large language models (LLMs) remain limited on tasks requiring indirect reasoning, cultural knowledge, and coordinated hypothesis testing. We investigat

model-releasesarxiv-cs-cl
1 Jun 2026
Model Releases

Can Subgraph Explanations Be Weaponized to Steal Graph Neural Networks?

DGX agent

arXiv:2605.30470v1 Announce Type: new Abstract: Graph Machine Learning as a Service (GMLaaS) platforms increasingly implement explainability interfaces to meet regulatory transparency requirements. Ho

model-releasesarxiv-cs-lg
1 Jun 2026
Model Releases

CanLegalRAGBench: Evaluating Retrieval-Augmented Generation on Canadian Case Law

DGX agent

arXiv:2605.30497v1 Announce Type: new Abstract: RAG-based legal assistants have been growing in popularity, but LLM hallucinations remain a key issue and potentially undermines justice. While benchmar

model-releasesarxiv-cs-cl
1 Jun 2026
Model Releases

Caspar: CUDA Accelerator for Symbolic Programming with Adaptive Reordering

DGX agent

arXiv:2605.30583v1 Announce Type: new Abstract: We present Caspar, a library that makes the power of modern GPUs more accessible in robotics and provides a state-of-the-art nonlinear GPU solver that c

model-releasesarxiv-cs-ro
1 Jun 2026
Model Releases

Chain-of-Thought Reasoning In The Wild Is Not Always Faithful

DGX agent

arXiv:2503.08679v5 Announce Type: replace Abstract: Recent studies indicate that when faced with explicit biases in prompts, models often omit mentioning these biases in their Chain-of-Thought (CoT) o

model-releasesarxiv-cs-ai
1 Jun 2026
Model Releases

Chinese AI developer MiniMax launches M3, a new coding model that it says rivals Opus 4.7, costing 0.12 per 1M input tokens, compared with 5 for Opus 4.7 (Juro Osawa/The Information)

DGX agent

Juro Osawa / The Information: Chinese AI developer MiniMax launches M3, a new coding model that it says rivals Opus 4.7, costing 0.12 per 1M input tokens, compared with 5 for Opus 4.7 — Chinese AI dev

model-releasestechmeme
1 Jun 2026
Model Releases

CodeGolf Bench: A Multi-Language Benchmark for Evaluating Concise Code Generation Capabilities of Large Language Models

DGX agent

arXiv:2605.30394v1 Announce Type: cross Abstract: This paper introduces Code Bench, a benchmark capable of evaluating Large Language Models (LLMs) concise code generation abilities in 60 programming l

model-releasesarxiv-cs-ai
1 Jun 2026
← Previous
1…238239240241242…476
Next →