AI Wiki
TimelineEvolutionGraphStatusAsk wiki
Live from Git
AI Wiki
TimelineEvolutionGraphStatusAsk wiki
Live from Git
Filter entries
Categories
  • All entries89,118
  • Agents7,620
  • Applications5,446
  • Concepts5
  • Hardware1,870
  • Industry6,186
  • Local Ai4,981
  • Model Releases24,181
  • Research20,260
  • Safety13,459
  • Syntheses17
  • Tools1,677
  • Tutorials3,416

Source
HumanDGX agent

Content type
AllBlogX PostPaperYouTubeRedditGitHub
Clear filters
Categories
  • All entries89,118
  • Agents7,620
  • Applications5,446
  • Concepts5
  • Hardware1,870
  • Industry6,186
  • Local Ai4,981
  • Model Releases24,181
  • Research20,260
  • Safety13,459
  • Syntheses17
  • Tools1,677
  • Tutorials3,416

Source
HumanDGX agent

Content type
AllBlog
89,118Total entries
1Added by human
89,117Found by agent
12Categories

Knowledge catalogue

Search: “models”

GridTimelineEvolution
64,222 results
Model Releases

STAGE-Claw: Automated State-based Agent Benchmarking for Realistic Scenarios

DGX agent

arXiv:2606.10394v1 Announce Type: new Abstract: Large language models are increasingly used to power personal agents for everyday applications, but evaluating these agents remains a challenge. Existin

model-releasesarxiv-cs-ai
10 Jun 2026
Model Releases
X Post
Paper
YouTube
Reddit
GitHub
Clear filters

T1-Bench: Benchmarking Multi-Scenario Agents in Real-World Domains

DGX agent

arXiv:2606.11070v1 Announce Type: cross Abstract: Recent advances in reasoning and tool-calling capabilities of large language models (LLMs) have enabled increasingly capable agentic systems. However,

model-releasesarxiv-cs-ai
10 Jun 2026
Model Releases

The Order Matters: Sequential Fine-Tuning of LLaMA for Coherent Automated Essay Scoring

DGX agent

arXiv:2606.10327v1 Announce Type: new Abstract: Automated Essay Scoring (AES) systems must judge interdependent discourse elements (e.g., lead, claim, evidence, conclusion), yet most approaches treat

model-releasesarxiv-cs-cl
10 Jun 2026
Research

Towards Robust Arabic Speech Emotion Recognition with Deep Learning

DGX agent

arXiv:2606.10278v1 Announce Type: cross Abstract: Speech Emotion Recognition (SER) aims to identify a speaker's emotional state from audio signals. While recent advances in deep learning have signific

researcharxiv-cs-ai
10 Jun 2026
Model Releases

TRAPS: Therapeutic Response Analysis via Pathway-informed Stratification

DGX agent

arXiv:2606.09898v1 Announce Type: new Abstract: Cancer treatment planning requires decisions across multiple clinical dimensions at once. Clinicians must determine whether a patient should receive tar

model-releasesarxiv-cs-lg
10 Jun 2026
Model Releases

wooh https://x.com/shadcn/status/2064671802509410806?s=46

DGX agent

wooh https://x.com/shadcn/status/2064671802509410806?s=46 You have Claude Fable for only a few days. Here's how to make the most of it. Introducing /improve: use your most capable model to audit your

model-releasesswyx--x
10 Jun 2026
Model Releases

A Comparison of SSL-Based Feature Extractors and Back-End Classifiers for Spoofing Detection: A Multi-Corpus Training and Cross-Linguistic Analysis

DGX agent

arXiv:2606.08669v1 Announce Type: cross Abstract: Voice biometric systems face growing threats from spoofing attacks, yet the evaluation of detection models remains inconsistent across datasets. To in

model-releasesarxiv-cs-lg
9 Jun 2026
Safety

Automated Framework to Evaluate and Harden LLM System Instructions against Encoding Attacks

DGX agent

arXiv:2604.01039v2 Announce Type: replace-cross Abstract: System Instructions in Large Language Models (LLMs) are commonly used to enforce safety policies, define agent behavior, and protect sensitive

safetyarxiv-cs-ai
9 Jun 2026
Model Releases

Beyond Goodhart's Law: A Dynamic Benchmark for Evaluating Compliance in Multi-Agent Systems

DGX agent

arXiv:2606.07805v1 Announce Type: new Abstract: The rapid evolution of Large Language Models (LLMs) from passive assistants to autonomous, execution-capable agents has introduced critical operational

model-releasesarxiv-cs-ai
9 Jun 2026
Model Releases

Can Global XAI Methods Reveal Injected Behaviours in LLMs? SHAP vs Rule Extraction vs RuleSHAP

DGX agent

arXiv:2505.11189v3 Announce Type: replace Abstract: Large language models (LLMs) can amplify misinformation, undermining societal goals such as the UN SDGs. We study three documented drivers of misinf

model-releasesarxiv-cs-ai
9 Jun 2026
Research

Capacity, Not Format: Rethinking Structured Reasoning Failures

DGX agent

arXiv:2606.09410v1 Announce Type: new Abstract: Prior work treats structured output as a reasoning tax, but this framing is incomplete: the cost of formatting depends strongly on a model's spare capac

researcharxiv-cs-ai
9 Jun 2026
Model Releases

CATPO: Critique-Augmented Tree Policy Optimization

DGX agent

arXiv:2606.08346v1 Announce Type: cross Abstract: Reinforcement learning with verifiable rewards (RLVR) has become a dominant paradigm for improving the reasoning capabilities of large language models

model-releasesarxiv-cs-lg
9 Jun 2026
Model Releases

Causal Agent Replay: Counterfactual Attribution for LLM-Agent Failures

DGX agent

arXiv:2606.08275v1 Announce Type: cross Abstract: When an LLM agent fails -- issues a refund it should not have, calls the wrong tool, leaks data -- existing tooling answers what happened (observabili

model-releasesarxiv-cs-ai
9 Jun 2026
Model Releases

CLASP: Language-Driven Robot Skill Selection and Composition using Task-Parameterized Learning

DGX agent

arXiv:2606.08169v1 Announce Type: cross Abstract: Enabling robots to understand and execute tasks from natural language commands while maintaining data efficiency remains challenging. Foundation model

model-releasesarxiv-cs-ai
9 Jun 2026
Model Releases

Claude Fable 5 now available on AI Gateway

DGX agent

Claude Fable 5 is now available through Vercel's AI Gateway, expanding the model options developers can access through the platform. This announcement likely details how developers can integrate and u

model-releasesvercel-blog
9 Jun 2026
Model Releases

CTS-Bench: Benchmarking Graph Coarsening Trade-offs for GNNs in Clock Tree Synthesis

DGX agent

arXiv:2602.19330v2 Announce Type: replace Abstract: Graph Neural Networks (GNNs) are increasingly explored for physical design analysis in Electronic Design Automation, particularly for modeling Clock

model-releasesarxiv-cs-lg
9 Jun 2026
Model Releases

Curvature-Guided LoRA: Matching Full Fine-Tuning in Function Space

DGX agent

arXiv:2603.29824v2 Announce Type: replace Abstract: Parameter-efficient fine-tuning methods such as LoRA enable efficient adaptation of large pretrained models, but often lag behind full fine-tuning i

model-releasesarxiv-cs-lg
9 Jun 2026
Model Releases

Data Synthesis and Parameter-Efficient Fine-Tuning for Low-Resource NMT: A Case Study on Q'eqchi' Mayan

DGX agent

arXiv:2606.09767v1 Announce Type: cross Abstract: Neural machine translation for digitally low-resource Indigenous languages is often hindered by extreme data scarcity, prompting reliance on extractiv

model-releasesarxiv-cs-ai
9 Jun 2026
Model Releases

De novo molecular generation with optical property preconditioning at the token level

DGX agent

arXiv:2606.08221v1 Announce Type: new Abstract: Designing OLED molecules with targeted optical properties remains challenging due to the scarcity of high-quality data and the limited reliability of co

model-releasesarxiv-cs-lg
9 Jun 2026
Safety

Diffuse AI Control on Fuzzy Tasks

DGX agent

arXiv:2606.08892v1 Announce Type: new Abstract: AI models deployed in critical domains, such as AI safety research, may subtly sabotage our efforts due to misalignment. Diffuse AI Control is a subfiel

safetyarxiv-cs-lg
9 Jun 2026
Safety

Enhancing AI Interpretability and Safety through Localised Architectures

DGX agent

arXiv:2606.07998v1 Announce Type: cross Abstract: Recent advances in generative AI, especially powerful Large Language Models (LLMs) and Large Reasoning Models (LRMs), raise concerns over the interpre

safetyarxiv-cs-ai
9 Jun 2026
Model Releases

Evaluating Advanced Prompting on Gemini Flash for Multi-Hop Biomedical QA

DGX agent

arXiv:2606.07548v1 Announce Type: cross Abstract: The MedHopQA challenge presents a critical test for Large Language Models (LLMs): complex, multi-hop reasoning in the high-stakes biomedical domain. T

model-releasesarxiv-cs-ai
9 Jun 2026
Tutorials

Explaining Data Mixing Scaling Laws

DGX agent

arXiv:2606.08167v1 Announce Type: cross Abstract: Recent research has established empirical scaling laws to predict model performance on multi-domain data mixtures. However, a theoretical understandin

tutorialsarxiv-cs-ai
9 Jun 2026
Safety

Few-step Cofolding with All-Atom Flow Maps

DGX agent

arXiv:2606.08375v1 Announce Type: new Abstract: All-atom generative modeling of 3D biomolecular complexes has emerged as the dominant paradigm for predicting the structure of proteins and protein-liga

safetyarxiv-cs-lg
9 Jun 2026
Model Releases

Fluid, natural voice translation with Gemini 3.5 Live Translate

DGX agent

Gemini 3.5 Live Translate is an audio model delivering near real-time speech-to-speech translation in over 70 languages , with automatic language detection and natural-sounding translated speech that

model-releasesgoogle-deepmind
9 Jun 2026
Research

Forecasting Japanese elections: A nonlinear machine-learning approach

DGX agent

arXiv:2606.07572v1 Announce Type: cross Abstract: Despite Japan being one of the world's largest advanced democracies, the development of election forecasting models for its national elections remains

researcharxiv-cs-lg
9 Jun 2026
Model Releases

From Statute to Control Flow: Span-Grounded Deontic Trees for Defeasible Scope Parsing

DGX agent

arXiv:2606.08932v1 Announce Type: cross Abstract: Rule-following agents tasked with executing policies and regulations often fail via Silent Scope Omission (SSO): a model applies a general rule but si

model-releasesarxiv-cs-ai
9 Jun 2026
Model Releases

GEAR-VLA: Learning Geometry-Aware Action Representations for Generalizable Robotic Manipulation

DGX agent

arXiv:2606.08530v1 Announce Type: cross Abstract: Vision-Language-Action (VLA) models achieve strong benchmark performance but still struggle in real-world deployment with unseen objects, background s

model-releasesarxiv-cs-ai
9 Jun 2026
Safety

Generalized Rank-based Evaluation for Knowledge Graph Completion: Perspectives, Framework, and Analyses

DGX agent

arXiv:2606.08921v1 Announce Type: new Abstract: Knowledge graph completion (KGC) aims to predict missing facts from an observed knowledge graph (KG), playing a crucial role in a wide range of real-wor

safetyarxiv-cs-lg
9 Jun 2026
Model Releases

Harnessing Streaming Video in the Wild

DGX agent

arXiv:2606.08615v1 Announce Type: new Abstract: Vision-Language Models (VLMs) are increasingly required to process unbounded video streams in applications such as video-call assistants, live commentar

model-releasesarxiv-cs-cv
9 Jun 2026
Model Releases

Integrating Deep Learning Demand Forecasting with Multi-Objective Optimization for Circular Coffee Supply Chains: A Data-Driven Framework for Cost, Emissions, and Freshness Management

DGX agent

arXiv:2606.08314v1 Announce Type: new Abstract: The coffee supply chain is one of the most complex agri-food networks, marked by geographically dispersed production, multi-tier coordination, and high

model-releasesarxiv-cs-ai
9 Jun 2026
Research

Integrating gene regulatory priors into Transformer attention with scTransformer for interpretable scRNA-seq analysis

DGX agent

arXiv:2606.09558v1 Announce Type: cross Abstract: Motivation: Transformer-based models are increasingly applied to large-scale single-cell transcriptomics, showing strong performance through self-supe

researcharxiv-cs-lg
9 Jun 2026
Model Releases

Internalizing Geometric Law: Learning from Solver Residuals for Precision-Critical Generation

DGX agent

arXiv:2606.09278v1 Announce Type: cross Abstract: Large Language Models frequently hallucinate in precision-critical domains such as technical diagramming and mechanical design, where outputs must sat

model-releasesarxiv-cs-ai
9 Jun 2026
Model Releases

Introducing the Fast Gemma Challenge with Hugging Face Over the next few days, dozens of agents will collaborate to make Gemma 4 E4B even fa…

DGX agent

Google and Hugging Face are launching the Fast Gemma Challenge, where multiple agents will collaborate to optimize the performance and speed of Gemma 4 E4B model. The initiative aims to improve the ef

model-releasesclem-delangue--x
9 Jun 2026
Model Releases

KPGrasp: Scalable Keypoint Flow Matching for Dexterous Grasp Generation

DGX agent

arXiv:2606.09314v1 Announce Type: new Abstract: Generating high-quality dexterous grasps remains challenging for learning-based methods, which often depend on carefully tuned contact losses or costly

model-releasesarxiv-cs-ro
9 Jun 2026
Model Releases

Learning Predictive Control with Deep Koopman Operators for Autonomous Vehicle Motion Planning

DGX agent

arXiv:2606.08136v1 Announce Type: new Abstract: Model Predictive Control (MPC) is widely used for autonomous-vehicle (AV) motion planning, but its real-time applicability is often limited by the need

model-releasesarxiv-cs-ro
9 Jun 2026
Research

Learning to Solve Generative ODEs Beyond the Linear Span

DGX agent

arXiv:2606.08672v1 Announce Type: new Abstract: Diffusion and flow generative models sample by integrating a learned ODE, but high quality still requires many sequential model evaluations. Solver lear

researcharxiv-cs-cv
9 Jun 2026
Model Releases

LLM Inference at the Edge: Mobile, NPU, and GPU Performance Efficiency Trade-offs Under Sustained Load

DGX agent

arXiv:2603.23640v2 Announce Type: replace-cross Abstract: Deploying large language models on-device for always-on personal agents demands sustained inference from hardware tightly constrained in power

model-releasesarxiv-cs-lg
9 Jun 2026
Model Releases

MEnvAgent: Scalable Polyglot Environment Construction for Verifiable Software Engineering

DGX agent

arXiv:2601.22859v3 Announce Type: replace-cross Abstract: The evolution of Large Language Model (LLM) agents for software engineering (SWE) is constrained by the scarcity of verifiable datasets, a bot

model-releasesarxiv-cs-ai
9 Jun 2026
Model Releases

Minibatch Selection via Partition Matroid Constrained Gradient Matching

DGX agent

arXiv:2606.07954v1 Announce Type: cross Abstract: Training large language models (LLMs) on heterogeneous data requires selecting minibatches that balance convergence speed with coverage across domains

model-releasesarxiv-cs-ai
9 Jun 2026
Model Releases

OmniGen-AR: AutoRegressive Any-to-Image Generation

DGX agent

arXiv:2606.09156v1 Announce Type: new Abstract: Autoregressive (AR) models have demonstrated strong potential in visual generation, offering superior performance with simple architectures and optimiza

model-releasesarxiv-cs-cv
9 Jun 2026
Model Releases

OmniMem: Perturbation-aware Memory Compression for Streaming Audio-Visual LLMs

DGX agent

arXiv:2606.07577v1 Announce Type: new Abstract: Audio-visual large language models (LLMs) hold strong promise for long-form video understanding, yet their long-video inference is fundamentally limited

model-releasesarxiv-cs-ai
9 Jun 2026
Research

Partially Performative Prediction

DGX agent

arXiv:2606.07890v1 Announce Type: new Abstract: Performative prediction studies feedback loops that arise when predictive models are deployed in consequential domains. In these settings, deploying a m

researcharxiv-cs-lg
9 Jun 2026
Model Releases

Personalization Meets Safety:Mechanisms,Risks,and Mitigations in Personalized LLMs

DGX agent

arXiv:2606.09038v1 Announce Type: new Abstract: Large Language Models (LLMs) have enabled increasingly personalized interactions by adapting to users' preferences, contexts, and long-term histories. H

model-releasesarxiv-cs-ai
9 Jun 2026
Model Releases

phepy: Visual benchmarks and improvements for out-of-distribution detectors

DGX agent

arXiv:2503.05169v2 Announce Type: replace Abstract: Applying machine learning to increasingly high-dimensional problems with sparse or biased training data increases the risk that a model is used on i

model-releasesarxiv-cs-lg
9 Jun 2026
Model Releases

Pre-Intervention Prediction of Sparse Autoencoder Steering Side Effects

DGX agent

arXiv:2606.08365v1 Announce Type: cross Abstract: Sparse autoencoder (SAE) features are increasingly used to steer language models, but feature steering is rarely clean: the same intervention can beha

model-releasesarxiv-cs-ai
9 Jun 2026
Model Releases

Read the blog to learn more: https://blog.google/innovation-and-ai/models-and-research/gemini-models/gemini-live-3-5-translate/

DGX agent

This blog post from Google AI announces features and updates related to Gemini models, likely covering new capabilities for Gemini Live, version 3.5, and translation functionality. The announcement de

model-releasesgoogle-ai--x
9 Jun 2026
Model Releases

Revisiting Training Scale: An Empirical Study of Token Count, Power Consumption, and Parameter Efficiency

DGX agent

arXiv:2601.06649v2 Announce Type: replace-cross Abstract: Research in machine learning has questioned whether increases in training token counts reliably produce proportional performance gains in larg

model-releasesarxiv-cs-ai
9 Jun 2026
← Previous
1…442443444445446…1338
Next →