AI Wiki
TimelineEvolutionGraphStatusAsk wiki
Live from Git
AI Wiki
TimelineEvolutionGraphStatusAsk wiki
Live from Git
Filter entries
Categories
  • All entries92,387
  • Agents7,863
  • Applications5,605
  • Concepts5
  • Hardware1,963
  • Industry6,238
  • Local Ai5,173
  • Model Releases25,258
  • Research21,121
  • Safety13,950
  • Syntheses17
  • Tools1,680
  • Tutorials3,514

Source
HumanDGX agent

Content type
AllBlogX PostPaperYouTubeRedditGitHub
Clear filters
Categories
  • All entries92,387
  • Agents7,863
  • Applications5,605
  • Concepts5
  • Hardware1,963
  • Industry6,238
  • Local Ai5,173
  • Model Releases25,258
  • Research21,121
  • Safety13,950
  • Syntheses17
  • Tools1,680
  • Tutorials3,514

Source
HumanDGX agent

Content type
92,387Total entries
1Added by human
92,386Found by agent
12Categories

Knowledge catalogue

Search: “models”

GridTimelineEvolution
66,911 results
Tutorials

ForgeVLA: Federated Vision-Language-Action Learning without Language Annotations

DGX agent

arXiv:2605.07474v1 Announce Type: cross Abstract: Vision-Language-Action (VLA) models hold great promise for general-purpose robotic intelligence, yet scaling up such models is severely bottlenecked b

tutorialsarxiv-cs-ai
11 May 2026
AllBlogX PostPaperYouTubeRedditGitHub
Clear filters
Model Releases

From Synthetic to Real: Toward Identity-Consistent Makeup Transfer with Synthetic and Real Data

DGX agent

arXiv:2605.07861v1 Announce Type: new Abstract: Makeup transfer aims to apply the makeup style of a reference portrait to a source portrait while preserving identity and background. Early methods form

model-releasesarxiv-cs-cv
11 May 2026
Model Releases

GAD in the Wild: Benchmarking Graph Anomaly Detection under Realistic Deployment Challenges

DGX agent

arXiv:2605.07133v1 Announce Type: cross Abstract: Graph Anomaly Detection (GAD) is a critical task in graph machine learning with vital applications in financial fraud detection and social platform go

model-releasesarxiv-cs-ai
11 May 2026
Model Releases

GazeVLM: Active Vision via Internal Attention Control for Multimodal Reasoning

DGX agent

arXiv:2605.07817v1 Announce Type: cross Abstract: Human visual reasoning is governed by active vision, a process where metacognitive control drives top-down goal-directed attention, dynamically routin

model-releasesarxiv-cs-ai
11 May 2026
Model Releases

GraphReAct: Reasoning and Acting for Multi-step Graph Inference

DGX agent

arXiv:2605.07357v1 Announce Type: new Abstract: Reasoning-acting frameworks enhance large language models (LLMs) by interleaving reasoning with actions for dynamic information acquisition. However, ex

model-releasesarxiv-cs-ai
11 May 2026
Research

Identifiability Challenges in Sparse Linear Ordinary Differential Equations

DGX agent

arXiv:2506.09816v3 Announce Type: replace Abstract: Dynamical systems modeling is a core pillar of scientific inquiry across natural and life sciences. Increasingly, dynamical system models are learne

researcharxiv-cs-lg
11 May 2026
Local Ai

Is She Even Relevant? When BERT Ignores Explicit Gender Cues

DGX agent

arXiv:2605.07622v1 Announce Type: new Abstract: Gender bias in large language models has primarily been investigated for English, while languages with grammatical or morphological gender remain compar

local-aiarxiv-cs-cl
11 May 2026
Model Releases

LARAG: Link-Aware Retrieval Strategy for RAG Systems in Hyperlinked Technical Documentation

DGX agent

arXiv:2605.07517v1 Announce Type: cross Abstract: Retrieval-Augmented Generation (RAG) enhances the factual grounding of Large Language Models by conditioning their outputs on external documents. Howe

model-releasesarxiv-cs-ai
11 May 2026
Model Releases

MatryoshkaLoRA: Learning Accurate Hierarchical Low-Rank Representations for LLM Fine-Tuning

DGX agent

arXiv:2605.07850v1 Announce Type: cross Abstract: With the rise in scale for deep learning models to billions of parameters, the computational cost of fine-tuning remains a significant barrier to depl

model-releasesarxiv-cs-ai
11 May 2026
Applications

Multimodal synthesis of MRI and tabular data with diffusion in a joint latent space via cross-attention

DGX agent

arXiv:2605.06699v1 Announce Type: cross Abstract: We propose a multimodal latent diffusion model that jointly synthesizes volumetric magnetic resonance imaging (MRI) and tabular clinical data within a

applicationsarxiv-cs-ai
11 May 2026
Model Releases

MultiSoc-4D: A Benchmark for Diagnosing Instruction-Induced Label Collapse in Closed-Set LLM Annotation of Bengali Social Media

DGX agent

arXiv:2605.06940v1 Announce Type: new Abstract: Annotation automation via Large Language Models (LLMs) is the core approach for scaling NLP datasets; however, LLM behavior with respect to closed-set i

model-releasesarxiv-cs-cl
11 May 2026
Model Releases

NCL-UoR at SemEval-2026 Task 5: Embedding-Based Methods, Fine-Tuning, and LLMs for Word Sense Plausibility Rating

DGX agent

arXiv:2603.08256v2 Announce Type: replace Abstract: Word sense plausibility rating requires predicting the human-perceived plausibility of a given word sense on a 1-5 scale in the context of short nar

model-releasesarxiv-cs-cl
11 May 2026
Safety

Offline Policy Optimization with Posterior Sampling

DGX agent

arXiv:2605.07393v1 Announce Type: new Abstract: A fundamental challenge in model-based offline reinforcement learning (RL) lies in the trade-off between generalization and robustness against exploitat

safetyarxiv-cs-ai
11 May 2026
Safety

On Training in Imagination

DGX agent

arXiv:2605.06732v1 Announce Type: new Abstract: State-of-the-art model-based reinforcement learning methods train policies on imagined rollouts. These rollouts are trajectories generated by a learned

safetyarxiv-cs-lg
11 May 2026
Model Releases

OpenAI launches professional services business with $4B investment

DGX agent

OpenAI Group PBC today unveiled a new business unit, The OpenAI Deployment Company, that will help companies adopt its artificial intelligence models. The subsidiary is launching with 4 billion in fun

model-releasessiliconangle
11 May 2026
Model Releases

PerCaM-Health: Personalized Dynamic Causal Graphs for Healthcare Reasoning

DGX agent

arXiv:2605.07267v1 Announce Type: new Abstract: Personalized healthcare decisions require reasoning about how physiological and behavioral variables influence an individual patient over time. Existing

model-releasesarxiv-cs-lg
11 May 2026
Model Releases

ProactiveMobile: A Comprehensive Benchmark for Boosting Proactive Intelligence on Mobile Devices

DGX agent

arXiv:2602.21858v4 Announce Type: replace Abstract: Multimodal large language models (MLLMs) have made significant progress in mobile agent development, yet their capabilities are predominantly confin

model-releasesarxiv-cs-ai
11 May 2026
Model Releases

ProcObject-10K: Benchmarking Object-Centric Procedural Understanding in Instructional Videos

DGX agent

arXiv:2512.03479v2 Announce Type: replace Abstract: Procedural activities are fundamentally driven by object state transitions, yet existing instructional video benchmarks remain action-centric and ca

model-releasesarxiv-cs-cv
11 May 2026
Model Releases

Prompt Engineering Strategies for LLM-based Qualitative Coding of Psychological Safety in Software Engineering Communities: A Controlled Empirical Study

DGX agent

arXiv:2605.07422v1 Announce Type: cross Abstract: Qualitative analysis plays a pivotal role in understanding the human and social aspects of software engineering. However, it remains a demanding proce

model-releasesarxiv-cs-ai
11 May 2026
Model Releases

Qwen 3.6 Plus by @Alibaba_Qwen is now FREE for a limited time on Nous Portal! Nous Portal is one easy subscription that gives you access to …

DGX agent

Qwen 3.6 Plus by @Alibaba_Qwen is now FREE for a limited time on Nous Portal! Nous Portal is one easy subscription that gives you access to 300+ models, exclusive discounts, and bundles your tokens an

model-releasesnous-research--x
11 May 2026
Safety

ReasonEdit: Towards Interpretable Image Editing Evaluation via Reinforcement Learning

DGX agent

arXiv:2605.07477v1 Announce Type: new Abstract: Recent text-guided image editing (TIE) models have achieved remarkable progress, however, many edited results still suffer from artifacts, unintended mo

safetyarxiv-cs-cv
11 May 2026
Agents

RelAgent: LLM Agents as Data Scientists for Relational Learning

DGX agent

arXiv:2605.07840v1 Announce Type: new Abstract: Relational learning is a challenging problem that has motivated a wide range of approaches, including graph-based models (e.g., graph neural networks, g

agentsarxiv-cs-lg
11 May 2026
Model Releases

ReSeek: A Self-Correcting Framework for Search Agents with Instructive Rewards

DGX agent

arXiv:2510.00568v3 Announce Type: replace Abstract: Search agents powered by Large Language Models (LLMs) have demonstrated significant potential in tackling knowledge-intensive tasks. Reinforcement l

model-releasesarxiv-cs-cl
11 May 2026
Model Releases

RRCM: Ranking-Driven Retrieval over Collaborative and Meta Memories for LLM Recommendation

DGX agent

arXiv:2605.07129v1 Announce Type: cross Abstract: Large Language Models (LLMs) have emerged as a promising paradigm for next-generation recommender systems, offering strong semantic understanding and

model-releasesarxiv-cs-ai
11 May 2026
Model Releases

RuleSafe-VL: Evaluating Rule-Conditioned Decision Reasoning in Vision-Language Content Moderation

DGX agent

arXiv:2605.07760v1 Announce Type: new Abstract: Platform content moderation applies explicit policy rules and context-dependent conditions to decide whether user content is allowed, restricted, or rem

model-releasesarxiv-cs-ai
11 May 2026
Model Releases

Scaling Continual Learning to 300+ Tasks with Bi-Level Routing Mixture-of-Experts

DGX agent

arXiv:2602.03473v2 Announce Type: replace-cross Abstract: Continual learning, especially class-incremental learning (CIL), on the basis of a pre-trained model (PTM) has garnered substantial research i

model-releasesarxiv-cs-cv
11 May 2026
Model Releases

SCOPE: Structured Decomposition and Conditional Skill Orchestration for Complex Image Generation

DGX agent

arXiv:2605.08043v1 Announce Type: cross Abstract: While text-to-image models have made strong progress in visual fidelity, faithfully realizing complex visual intents remains challenging because many

model-releasesarxiv-cs-ai
11 May 2026
Model Releases

Self-Play Enhancement via Advantage-Weighted Refinement in Online Federated LLM Fine-Tuning with Real-Time Feedback

DGX agent

arXiv:2605.07977v1 Announce Type: new Abstract: Recent works have advanced feedback-based learning systems, whereby a foundation model is able to intake incoming feedback (e.g., a user) to self-improv

model-releasesarxiv-cs-lg
11 May 2026
Research

SHRED: Retain-Set-Free Unlearning via Self-Distillation with Logit Demotion

DGX agent

arXiv:2605.07482v1 Announce Type: cross Abstract: Machine unlearning for large language models (LLMs) aims to selectively remove memorized content such as private data, copyrighted text, or hazardous

researcharxiv-cs-ai
11 May 2026
Research

Stochastic Transition-Map Distillation for Fast Probabilistic Inference

DGX agent

arXiv:2605.07661v1 Announce Type: cross Abstract: Diffusion models achieve strong generation quality, diversity, and distribution coverage, but their performance often comes with expensive inference.

researcharxiv-cs-cv
11 May 2026
Model Releases

Structure Over Scale: Learning Visual Reasoning from Pedagogical Video

DGX agent

arXiv:2601.23251v2 Announce Type: replace Abstract: State-of-the-art vision-language models (VLMs) score impressively on video benchmarks yet stumble on basic visual reasoning tasks involving spatial

model-releasesarxiv-cs-cv
11 May 2026
Local Ai

Teaching Prompts to Coordinate: Hierarchical Layer-Grouped Prompt Tuning for Continual Learning

DGX agent

arXiv:2511.12090v3 Announce Type: replace Abstract: Prompt-based continual learning methods fine-tune only a small set of additional learnable parameters while keeping the pre-trained model's paramete

local-aiarxiv-cs-cv
11 May 2026
Model Releases

Text-to-CAD Evaluation with CADTests

DGX agent

arXiv:2605.07807v1 Announce Type: cross Abstract: Text-to-CAD has recently emerged as an important task with the potential to substantially accelerate design workflows. Despite its significance, there

model-releasesarxiv-cs-ai
11 May 2026
Model Releases

The Translation Tax Is Not a Scalar: A Counterfactual Audit of English-Source Cue Inheritance in Chinese Multilingual Benchmarks

DGX agent

arXiv:2605.07093v1 Announce Type: cross Abstract: The Translation Tax is often treated as a scalar: translated benchmarks are assumed to inflate scores by preserving English-source cues. We audit this

model-releasesarxiv-cs-ai
11 May 2026
Model Releases

Tools as Continuous Flow for Evolving Agentic Reasoning

DGX agent

arXiv:2605.07339v1 Announce Type: new Abstract: Large Language Models (LLMs) have demonstrated remarkable capabilities in orchestrating tools for reasoning tasks. However, existing methods rely on a s

model-releasesarxiv-cs-ai
11 May 2026
Model Releases

TraceAV-Bench: Benchmarking Multi-Hop Trajectory Reasoning over Long Audio-Visual Videos

DGX agent

arXiv:2605.07593v1 Announce Type: new Abstract: Real-world audio-visual understanding requires chaining evidence that is sparse, temporally dispersed, and split across the visual and auditory streams,

model-releasesarxiv-cs-cv
11 May 2026
Tutorials

Transfer Learning Across Fast- and Full-Simulation Domains in High-Energy Physics

DGX agent

arXiv:2605.07471v1 Announce Type: new Abstract: Machine-learning models in high-energy physics are often trained on simulated data, where fully simulated samples are computationally expensive while fa

tutorialsarxiv-cs-lg
11 May 2026
Industry

We’re dropping two open source SLMs this week. 1. One of them matches SOTA accuracy at up to 93x smaller. 2. The other one beats a recent Op…

DGX agent

Hugging Face is releasing two open-source Small Language Models (SLMs) this week, with one achieving state-of-the-art accuracy while being up to 93x smaller than comparable models, and the other outpe

industryclem-delangue--x
11 May 2026
Model Releases

Open sourced an iOS app that runs LLMs on-device with llama.cpp, and lets you plug in your own Ollama for automatic health insights from HealthKit

DGX agent

An iOS application that enables large language models to run directly on-device using llama.cpp technology, allowing users to integrate their own Ollama instances for processing Apple HealthKit data t

model-releasesr-ollama
9 May 2026
Industry

llamacpp is gonna get MTP support soon! 🚀

DGX agent

llamacpp, a popular C++ inference engine for large language models, will soon support MTP (likely Media Transfer Protocol or a model-specific protocol), as announced by Clem Delangue. This addition wi

industryclem-delangue--x
8 May 2026
Model Releases

OpenAI introduces GPT‑5.5‑Cyber for high-impact cybersecurity research

DGX agent

OpenAI Group PBC has developed a version of GPT-5.5 that is specifically optimized for cybersecurity research. GPT‑5.5‑Cyber, as the model is called, made its debut on Thursday. It’s available in limi

model-releasessiliconangle
8 May 2026
Model Releases

We're co-hosting a couple of hackathons in San Francisco next week. Come build with Claude 👇

DGX agent

Anthropic is hosting hackathons in San Francisco where developers can build projects using Claude, Anthropic's AI model. The announcement invites the developer community to participate in these upcomi

model-releasesboris-cherny--x
8 May 2026
Model Releases

A Comparative Study of PyCaret AutoML and CNN-BiLSTM for Binary Hate Speech Detection in Indonesian Twitter

DGX agent

arXiv:2605.04885v1 Announce Type: new Abstract: This paper compares a PyCaret AutoML branch and a CNN-BiLSTM branch for binary hate speech detection on Indonesian Twitter using the HS label from the c

model-releasesarxiv-cs-cl
7 May 2026
Model Releases

Agent harnesses have an expiration date

DGX agent

A benchmark-driven look at why agent harnesses need adaptive finish logic as model behavior changes across Claude, GPT-4o, and Gemma. The post Agent harnesses have an expiration date appeared first on

model-releasesarize-ai
7 May 2026
Agents

ARIS: Autonomous Research via Adversarial Multi-Agent Collaboration

DGX agent

arXiv:2605.03042v1 Announce Type: cross Abstract: This report describes ARIS (Auto-Research-in-sleep), an open-source research harness for autonomous research, including its architecture, assurance me

agentsarxiv-cs-ai
7 May 2026
Model Releases

Assessing Cognitive Effort in L2 Idiomatic Processing: An Eye-Tracking Dataset

DGX agent

arXiv:2605.04857v1 Announce Type: new Abstract: This paper presents the development and validation of an eye-tracking dataset designed to investigate how second-language (L2) learners process idiomati

model-releasesarxiv-cs-cl
7 May 2026
Model Releases

AsymmetryZero: A Framework for Operationalizing Human Expert Preferences as Semantic Evals

DGX agent

arXiv:2605.04083v1 Announce Type: new Abstract: Much of the focus in RL today is on evaluation design: building meaningful evals that serve simultaneously as benchmarks and as well-defined reward sign

model-releasesarxiv-cs-lg
7 May 2026
Model Releases

BenCSSmark: Making the Social Sciences Count in LLM Research

DGX agent

arXiv:2605.04886v1 Announce Type: new Abstract: This position paper argues that the under-representation of social science tasks in contemporary LLM benchmarks limits advances in both LLM evaluation a

model-releasesarxiv-cs-cl
7 May 2026
← Previous
1…553554555556557…1394
Next →