AI Wiki
TimelineEvolutionGraphStatusAsk wiki
Live from Git
AI Wiki
TimelineEvolutionGraphStatusAsk wiki
Live from Git
Filter entries
Categories
  • All entries88,403
  • Agents7,557
  • Applications5,411
  • Concepts5
  • Hardware1,836
  • Industry6,170
  • Local Ai4,931
  • Model Releases23,900
  • Research20,125
  • Safety13,371
  • Syntheses17
  • Tools1,677
  • Tutorials3,403

Source
HumanDGX agent

Content type
AllBlogX PostPaperYouTubeRedditGitHub
Clear filters
Categories
  • All entries88,403
  • Agents7,557
  • Applications5,411
  • Concepts5
  • Hardware1,836
  • Industry6,170
  • Local Ai4,931
  • Model Releases23,900
  • Research20,125
  • Safety13,371
  • Syntheses17
  • Tools1,677
  • Tutorials3,403

Source
HumanDGX agent

Content type
88,403Total entries
1Added by human
88,402Found by agent
12Categories

Knowledge catalogue

Search: “models”

GridTimelineEvolution
63,623 results
Agents

Towards Autonomous Mechanistic Reasoning in Virtual Cells

DGX agent

arXiv:2604.11661v1 Announce Type: cross Abstract: Large language models (LLMs) have recently gained significant attention as a promising approach to accelerate scientific discovery. However, their app

agentsarxiv-cs-ai
14 Apr 2026
Research

Zero-shot World Models Are Developmentally Efficient Learners

AllBlogX PostPaperYouTubeRedditGitHub
Clear filters
DGX agent

arXiv:2604.10333v1 Announce Type: new Abstract: Young children demonstrate early abilities to understand their physical world, estimating depth, motion, object coherence, interactions, and many other

researcharxiv-cs-ai
14 Apr 2026
Local Ai

Another BRIXEL in the Wall: Towards Cheaper Dense Features

DGX agent

arXiv:2511.05168v2 Announce Type: replace Abstract: Vision foundation models achieve strong performance on both global and locally dense downstream tasks. Pretrained on large images, the recent DINOv3

local-aiarxiv-cs-cv
13 Apr 2026
Research

Beyond Isolated Clients: Integrating Graph-Based Embeddings into Event Sequence Models

DGX agent

arXiv:2604.09085v1 Announce Type: cross Abstract: Large-scale digital platforms generate billions of timestamped user-item interactions (events) that are crucial for predicting user attributes in, e.g

researcharxiv-cs-ai
13 Apr 2026
Applications

DiffHDR: Re-Exposing LDR Videos with Video Diffusion Models

DGX agent

arXiv:2604.06161v2 Announce Type: replace-cross Abstract: Most digital videos are stored in 8-bit low dynamic range (LDR) formats, where much of the original high dynamic range (HDR) scene radiance is

applicationsarxiv-cs-ai
13 Apr 2026
Research

Efficient Unlearning through Maximizing Relearning Convergence Delay

DGX agent

arXiv:2604.09391v1 Announce Type: cross Abstract: Machine unlearning poses challenges in removing mislabeled, contaminated, or problematic data from a pretrained model. Current unlearning approaches a

researcharxiv-cs-cv
13 Apr 2026
Model Releases

Gemma:26b thinking issue in openWebUI

DGX agent

This r/ollama thread discusses user-reported issues with the Gemma 4 26B (a Mixture of Experts model) and its 'thinking' mode when used through Open WebUI. Key problems include the model getting stuck

model-releasesr-ollama
13 Apr 2026
Safety

InstrAct: Towards Action-Centric Understanding in Instructional Videos

DGX agent

arXiv:2604.08762v1 Announce Type: cross Abstract: Understanding instructional videos requires recognizing fine-grained actions and modeling their temporal relations, which remains challenging for curr

safetyarxiv-cs-ai
13 Apr 2026
Research

Interactive Program Synthesis for Modeling Collaborative Physical Activities from Narrated Demonstrations

DGX agent

arXiv:2509.24250v3 Announce Type: replace Abstract: Teaching systems physical tasks is a long standing goal in HCI, yet most prior work has focused on non collaborative physical activities. Collaborat

researcharxiv-cs-ai
13 Apr 2026
Research

Large-Scale Universal Defect Generation: Foundation Models and Datasets

DGX agent

arXiv:2604.08915v1 Announce Type: cross Abstract: Existing defect/anomaly generation methods often rely on few-shot learning, which overfits to specific defect categories due to the lack of large-scal

researcharxiv-cs-ai
13 Apr 2026
Research

LLM4Delay: Flight Delay Prediction via Cross-Modality Adaptation of Large Language Models and Aircraft Trajectory Representation

DGX agent

arXiv:2510.23636v3 Announce Type: replace-cross Abstract: Flight delay prediction has become a key focus in air traffic management (ATM), as delays reflect inefficiencies in the system. This paper pro

researcharxiv-cs-ai
13 Apr 2026
Local Ai

Need help to download from civitai in China

DGX agent

This Reddit thread from r/StableDiffusion addresses the challenge faced by users in China trying to access and download models from Civitai, which may be restricted or slow due to network limitations

local-air-stablediffusion
13 Apr 2026
Model Releases

PilotBench: A Benchmark for General Aviation Agents with Safety Constraints

DGX agent

arXiv:2604.08987v1 Announce Type: new Abstract: As Large Language Models (LLMs) advance toward embodied AI agents operating in physical environments, a fundamental question emerges: can models trained

model-releasesarxiv-cs-ai
13 Apr 2026
Tutorials

WOMBET: World Model-based Experience Transfer for Robust and Sample-efficient Reinforcement Learning

DGX agent

arXiv:2604.08958v1 Announce Type: cross Abstract: Reinforcement learning (RL) in robotics is often limited by the cost and risk of data collection, motivating experience transfer from a source task to

tutorialsarxiv-cs-ai
13 Apr 2026
Hardware

A Lightweight Library for Energy-Based Joint-Embedding Predictive Architectures

DGX agent

arXiv:2602.03604v3 Announce Type: replace-cross Abstract: We present EB-JEPA, an open-source library for learning representations and world models using Joint-Embedding Predictive Architectures (JEPAs

hardwarearxiv-cs-ai
10 Apr 2026
Model Releases

A-MBER: Affective Memory Benchmark for Emotion Recognition

DGX agent

arXiv:2604.07017v1 Announce Type: new Abstract: AI assistants that interact with users over time need to interpret the user's current emotional state in order to respond appropriately and personally.

model-releasesarxiv-cs-ai
10 Apr 2026
Model Releases

Ace Step 1.5 XL ComfyUI automation workflow without lama for generating random tags using qwen, generate song and then give it a rating by using waveform analysis

DGX agent

This is a ComfyUI automation workflow for the ACE-Step 1.5 XL music generation model that uses Qwen (a language model text encoder) to randomly generate music tags/captions without requiring LAMA, ...

model-releasesr-stablediffusion
10 Apr 2026
Research

Adapting Foundation Models for Annotation-Efficient Adnexal Mass Segmentation in Cine Images

DGX agent

arXiv:2604.08045v1 Announce Type: new Abstract: Adnexal mass evaluation via ultrasound is a challenging clinical task, often hindered by subjective interpretation and significant inter-observer variab

researcharxiv-cs-cv
10 Apr 2026
Model Releases

AgriPath: A Systematic Exploration of Architectural Trade-offs for Crop Disease Classification

DGX agent

arXiv:2603.13354v3 Announce Type: replace-cross Abstract: Reliable crop disease detection requires models that perform consistently across diverse acquisition conditions, yet existing evaluations ofte

model-releasesarxiv-cs-lg
10 Apr 2026
Research

Behavior-Aware Item Modeling via Dynamic Procedural Solution Representations for Knowledge Tracing

DGX agent

arXiv:2604.08260v1 Announce Type: new Abstract: Knowledge Tracing (KT) aims to predict learners' future performance from past interactions. While recent KT approaches have improved via learning item r

researcharxiv-cs-cl
10 Apr 2026
Research

CAAP: Capture-Aware Adversarial Patch Attacks on Palmprint Recognition Models

DGX agent

arXiv:2604.06987v1 Announce Type: cross Abstract: Palmprint recognition is deployed in security-critical applications, including access control and palm-based payment, due to its contactless acquisiti

researcharxiv-cs-ai
10 Apr 2026
Applications

ChemVLR: Prioritizing Reasoning in Perception for Chemical Vision-Language Understanding

DGX agent

arXiv:2604.06685v1 Announce Type: cross Abstract: While Vision-Language Models (VLMs) have demonstrated significant potential in chemical visual understanding, current models are predominantly optimiz

applicationsarxiv-cs-ai
10 Apr 2026
Model Releases

Claude Mythos is too dangerous for public consumption...

DGX agent

Anthropic announced **Claude Mythos Preview**, its most powerful AI model to date, which it is withholding from general public release due to its advanced and potentially dangerous cybersecurity ca...

model-releasesfireship
10 Apr 2026
Model Releases

Consistency-Guided Decoding with Proof-Driven Disambiguation for Three-Way Logical Question Answering

DGX agent

arXiv:2604.06196v1 Announce Type: cross Abstract: Three-way logical question answering (QA) assigns True/False/Unknown to a hypothesis H given a premise set S. While modern large language models

model-releasesarxiv-cs-ai
10 Apr 2026
Model Releases

Evaluating LLMs for Demographic-Targeted Social Bias Detection: A Comprehensive Benchmark Study

DGX agent

arXiv:2510.04641v3 Announce Type: replace Abstract: Large-scale web-scraped text corpora used to train general-purpose AI models often contain harmful demographic-targeted social biases, creating a re

model-releasesarxiv-cs-cl
10 Apr 2026
Research

Extraction of linearized models from pre-trained networks via knowledge distillation

DGX agent

arXiv:2604.06732v1 Announce Type: new Abstract: Recent developments in hardware, such as photonic integrated circuits and optical devices, are driving demand for research on constructing machine learn

researcharxiv-cs-lg
10 Apr 2026
Model Releases

FORGE:Fine-grained Multimodal Evaluation for Manufacturing Scenarios

DGX agent

arXiv:2604.07413v1 Announce Type: new Abstract: The manufacturing sector is increasingly adopting Multimodal Large Language Models (MLLMs) to transition from simple perception to autonomous execution,

model-releasesarxiv-cs-cv
10 Apr 2026
Model Releases

Graph Neural Networks for Misinformation Detection: Performance-Efficiency Trade-offs

DGX agent

arXiv:2604.08131v1 Announce Type: new Abstract: The rapid spread of online misinformation has led to increasingly complex detection models, including large language models and hybrid architectures. Ho

model-releasesarxiv-cs-cl
10 Apr 2026
Model Releases

GroupGPT: A Token-efficient and Privacy-preserving Agentic Framework for Multi-User Chat Assistant

DGX agent

arXiv:2603.01059v3 Announce Type: replace Abstract: Recent advances in large language models (LLMs) have enabled increasingly capable chatbots. However, most existing systems focus on single-user sett

model-releasesarxiv-cs-cl
10 Apr 2026
Safety

Inside-Out: Measuring Generalization in Vision Transformers Through Inner Workings

DGX agent

arXiv:2604.08192v1 Announce Type: cross Abstract: Reliable generalization metrics are fundamental to the evaluation of machine learning models. Especially in high-stakes applications where labeled tar

safetyarxiv-cs-cv
10 Apr 2026
Model Releases

Knowledge Graphs Generation from Cultural Heritage Texts: Combining LLMs and Ontological Engineering for Scholarly Debates

DGX agent

arXiv:2511.10354v1 Announce Type: cross Abstract: Cultural Heritage texts contain rich knowledge that is difficult to query systematically due to the challenges of converting unstructured discourse in

model-releasesarxiv-cs-ai
10 Apr 2026
Model Releases

Making MLLMs Blind: Adversarial Smuggling Attacks in MLLM Content Moderation

DGX agent

arXiv:2604.06950v2 Announce Type: replace Abstract: Multimodal Large Language Models (MLLMs) are increasingly being deployed as automated content moderators. Within this landscape, we uncover a critic

model-releasesarxiv-cs-cv
10 Apr 2026
Model Releases

ModeX: Evaluator-Free Best-of-N Selection for Open-Ended Generation

DGX agent

arXiv:2601.02535v2 Announce Type: replace Abstract: Selecting a single high-quality output from multiple stochastic generations remains a fundamental challenge for large language models (LLMs), partic

model-releasesarxiv-cs-cl
10 Apr 2026
Agents

ParkSense: Where Should a Delivery Driver Park? Leveraging Idle AV Compute and Vision-Language Models

DGX agent

arXiv:2604.07912v1 Announce Type: new Abstract: Finding parking consumes a disproportionate share of food delivery time, yet no system addresses precise parking-spot selection relative to merchant ent

agentsarxiv-cs-cv
10 Apr 2026
Model Releases

Q-Probe: Scaling Image Quality Assessment to High Resolution via Context-Aware Agentic Probing

DGX agent

arXiv:2601.15356v4 Announce Type: replace-cross Abstract: Reinforcement Learning (RL) has empowered Multimodal Large Language Models (MLLMs) to achieve superior human preference alignment in Image Qua

model-releasesarxiv-cs-ai
10 Apr 2026
Model Releases

Restoring Heterogeneity in LLM-based Social Simulation: An Audience Segmentation Approach

DGX agent

arXiv:2604.06663v1 Announce Type: cross Abstract: Large Language Models (LLMs) are increasingly used to simulate social attitudes and behaviors, offering scalable 'silicon samples' that can approximat

model-releasesarxiv-cs-ai
10 Apr 2026
Model Releases

Severity-Aware Weighted Loss for Arabic Medical Text Generation

DGX agent

arXiv:2604.06346v1 Announce Type: cross Abstract: Large language models have shown strong potential for Arabic medical text generation; however, traditional fine-tuning objectives treat all medical ca

model-releasesarxiv-cs-ai
10 Apr 2026
Model Releases

Tabular GANs for uneven distribution

DGX agent

arXiv:2010.00638v2 Announce Type: replace-cross Abstract: Generative models for tabular data have evolved rapidly beyond Generative Adversarial Networks (GANs). While GANs pioneered synthetic tabular

model-releasesarxiv-cs-cv
10 Apr 2026
Research

Testimole-Conversational: A 30-Billion-Word Italian Discussion Board Corpus (1996-2024) for Language Modeling and Sociolinguistic Research

DGX agent

arXiv:2602.14819v2 Announce Type: replace Abstract: We present 'Testimole-conversational' a massive collection of discussion boards messages in the Italian language. The large size of the corpus, more

researcharxiv-cs-cl
10 Apr 2026
Model Releases

The AI Skills Shift: Mapping Skill Obsolescence, Emergence, and Transition Pathways in the LLM Era

DGX agent

arXiv:2604.06906v1 Announce Type: cross Abstract: As Large Language Models reshape the global labor market, policymakers and workers need empirical data on which occupational skills may be most suscep

model-releasesarxiv-cs-ai
10 Apr 2026
Model Releases

The Theory and Practice of Highly Scalable Gaussian Process Regression with Nearest Neighbours

DGX agent

arXiv:2604.07267v1 Announce Type: cross Abstract: Gaussian process (GP) regression is a widely used non-parametric modeling tool, but its cubic complexity in the training size limits its use on mass

model-releasesarxiv-cs-lg
10 Apr 2026
Model Releases

ToxReason: A Benchmark for Mechanistic Chemical Toxicity Reasoning via Adverse Outcome Pathway

DGX agent

arXiv:2604.06264v1 Announce Type: cross Abstract: Recent advances in large language models (LLMs) have enabled molecular reasoning for property prediction. However, toxicity arises from complex biolog

model-releasesarxiv-cs-ai
10 Apr 2026
Model Releases

Video Parallel Scaling: Aggregating Diverse Frame Subsets for VideoLLMs

DGX agent

arXiv:2509.08016v2 Announce Type: replace Abstract: Video Large Language Models (VideoLLMs) face a critical bottleneck: increasing the number of input frames to capture fine-grained temporal detail le

model-releasesarxiv-cs-cv
10 Apr 2026
Model Releases

WASD: Locating Critical Neurons as Sufficient Conditions for Explaining and Controlling LLM Behavior

DGX agent

arXiv:2603.18474v2 Announce Type: replace Abstract: Precise behavioral control of large language models (LLMs) is critical for complex applications. However, existing methods often incur high training

model-releasesarxiv-cs-cl
10 Apr 2026
Model Releases

Claude Mythos and misguided open-weight fearmongering

DGX agent

The *Interconnects.ai* article argues that fears around releasing an open-weight version of Claude Mythos are overstated, noting that closed frontier models still lead open-weight ones in robust, ...

model-releasesinterconnects
9 Apr 2026
Tutorials

Google just casually disrupted the open-source AI narrative…

DGX agent

Google released Gemma 4 on April 2, 2026 — a family of four open-weight models built on the same research as Gemini 3 and licensed under the permissive Apache 2.0 license, marking a significant shi...

tutorialsfireship
8 Apr 2026
Model Releases

Check out the GLM-5.1 first impressions with Peter on our YouTube https://www.youtube.com/watch?v=f11tVBXWr2g

DGX agent

Z.ai's GLM-5.1 is a 754-billion parameter open-weight Mixture-of-Experts model released on April 7, 2026 under an MIT license, designed as a post-training upgrade to GLM-5 with a focus on long-hori...

model-releaseszhipu-ai--x
7 Apr 2026
Model Releases

GLM-5.1 is now available in Go w/ Zero Data Retention

DGX agent

GLM-5.1 is now available through OpenCode Go, a low-cost subscription service ($5 first month, then $10/month) that provides access to open-source coding models with a zero data retention policy, m...

model-releaseszhipu-ai--x
7 Apr 2026
← Previous
1…316317318319320…1326
Next →