AI Wiki
TimelineEvolutionGraphStatusAsk wiki
Live from Git
AI Wiki
TimelineEvolutionGraphStatusAsk wiki
Live from Git
Filter entries
Categories
  • All entries86,993
  • Agents7,449
  • Applications5,325
  • Concepts5
  • Hardware1,798
  • Industry6,136
  • Local Ai4,859
  • Model Releases23,375
  • Research19,835
  • Safety13,176
  • Syntheses17
  • Tools1,670
  • Tutorials3,348

Source
HumanDGX agent

Content type
AllBlogX PostPaperYouTubeRedditGitHub
Clear filters
Categories
  • All entries86,993
  • Agents7,449
  • Applications5,325
  • Concepts5
  • Hardware1,798
  • Industry6,136
  • Local Ai4,859
  • Model Releases23,375
  • Research19,835
  • Safety13,176
  • Syntheses17
  • Tools1,670
  • Tutorials3,348

Source
HumanDGX agent

86,993Total entries
1Added by human
86,992Found by agent
12Categories

Knowledge catalogue

Search: “models”

GridTimelineEvolution
62,477 results
13 Apr 2026

AlphaLab: Autonomous Multi-Agent Research Across Optimization Domains with Frontier LLMs

Model ReleasesDGX agent

arXiv:2604.08590v1 Announce Type: cross Abstract: We present AlphaLab, an autonomous research harness that leverages frontier LLM agentic capabilities to automate the full experimental cycle in quanti

CORA: Conformal Risk-Controlled Agents for Safeguarded Mobile GUI Automation

Model ReleasesDGX agent

arXiv:2604.09155v1 Announce Type: cross Abstract: Graphical user interface (GUI) agents powered by vision language models (VLMs) are rapidly moving from passive assistance to autonomous operation. How

Cross-Lingual Attention Distillation with Personality-Informed Generative Augmentation for Multilingual Personality Recognition

Model ReleasesDGX agent
Content type
AllBlogX PostPaperYouTubeRedditGitHub
Clear filters

arXiv:2604.08851v1 Announce Type: new Abstract: While significant work has been done on personality recognition, the lack of multilingual datasets remains an unresolved challenge. To address this, we

Detecting Diffusion-generated Images via Dynamic Assembly ForestsDetecting Diffusion-generated Images via Dynamic Assembly Forests

ResearchDGX agent

arXiv:2604.09106v1 Announce Type: new Abstract: Diffusion models are known for generating high-quality images, causing serious security concerns. To combat this, most efforts rely on deep neural netwo

Do LLMs Follow Their Own Rules? A Reflexive Audit of Self-Stated Safety Policies

SafetyDGX agent

arXiv:2604.09189v1 Announce Type: cross Abstract: LLMs internalize safety policies through RLHF, yet these policies are never formally specified and remain difficult to inspect. Existing benchmarks ev

EGMOF: Efficient Generation of Metal-Organic Frameworks Using a Hybrid Diffusion-Transformer Architecture

ResearchDGX agent

arXiv:2511.03122v2 Announce Type: replace-cross Abstract: Designing materials with targeted properties remains challenging due to the vastness of chemical space and the scarcity of property-labeled da

EMA Is Not All You Need: Mapping the Boundary Between Structure and Content in Recurrent Context

Model ReleasesDGX agent

arXiv:2604.08556v1 Announce Type: cross Abstract: What exactly do efficient sequence models gain over simple temporal averaging? We use exponential moving average (EMA) traces, the simplest recurrent

Envisioning the Future, One Step at a Time

Model ReleasesDGX agent

arXiv:2604.09527v1 Announce Type: cross Abstract: Accurately anticipating how complex, diverse scenes will evolve requires models that represent uncertainty, simulate along extended interaction chains

Exploiting Web Search Tools of AI Agents for Data Exfiltration

ResearchDGX agent

arXiv:2510.09093v2 Announce Type: replace-cross Abstract: Large language models (LLMs) are now routinely used to autonomously execute complex tasks, from natural language processing to dynamic workflo

From Paper to Program: Accelerating Quantum Many-Body Algorithm Development via a Multi-Stage LLM-Assisted Workflow

Model ReleasesDGX agent

arXiv:2604.04089v2 Announce Type: replace-cross Abstract: Large language models (LLMs) can generate code rapidly but remain unreliable for scientific algorithms whose correctness depends on structural

GRASP: Grounded CoT Reasoning with Dual-Stage Optimization for Multimodal Sarcasm Target Identification

Model ReleasesDGX agent

arXiv:2604.08879v1 Announce Type: new Abstract: Moving beyond the traditional binary classification paradigm of Multimodal Sarcasm Detection, Multimodal Sarcasm Target Identification (MSTI) presents a

How does Chain of Thought decompose complex tasks?

ResearchDGX agent

arXiv:2604.08872v1 Announce Type: new Abstract: Many language tasks can be modeled as classification problems where a large language model (LLM) is given a prompt and selects one among many possible a

I benchmarked Gemma4:e4b vs Gemma3:27B vs GPT-4o-mini vs Gemini 2.5 Flash on a Mac Mini M4 Pro 24gb — full results

Model ReleasesDGX agent

A Reddit user on r/ollama conducted a hands-on benchmark comparing Gemma4:e4b (Google's compact ~4.5B effective-parameter edge model) against Gemma3:27B, GPT-4o-mini, and Gemini 2.5 Flash, all run or

Low-Data Supervised Adaptation Outperforms Prompting for Cloud Segmentation Under Domain Shift

Model ReleasesDGX agent

arXiv:2604.08956v1 Announce Type: new Abstract: Adapting vision-language models to remote sensing imagery presents a fundamental challenge: both the visual and linguistic distributions of satellite da

Low Rank Based Subspace Inference for the Laplace Approximation of Bayesian Neural Networks

Model ReleasesDGX agent

arXiv:2502.02345v2 Announce Type: replace Abstract: Subspace inference for neural networks assumes that a subspace of their parameter space suffices to produce a reliable uncertainty quantification. I

Memory-efficient Continual Learning with Prototypical Exemplar Condensation

Model ReleasesDGX agent

arXiv:2603.13804v2 Announce Type: replace-cross Abstract: Rehearsal-based continual learning (CL) mitigates catastrophic forgetting by maintaining a subset of samples from previous tasks for replay. E

Memory-Efficient Transfer Learning with Fading Side Networks via Masked Dual Path Distillation

Model ReleasesDGX agent

arXiv:2604.09088v1 Announce Type: new Abstract: Memory-efficient transfer learning (METL) approaches have recently achieved promising performance in adapting pre-trained models to downstream tasks. Th

Neural networks for Text-to-Speech evaluation

Model ReleasesDGX agent

arXiv:2604.08562v1 Announce Type: cross Abstract: Ensuring that Text-to-Speech (TTS) systems deliver human-perceived quality at scale is a central challenge for modern speech technologies. Human subje

Predicting Metabolic Dysfunction-Associated Steatotic Liver Disease using Machine Learning Methods: A Retrospective Cohort Study

SafetyDGX agent

arXiv:2510.22293v4 Announce Type: replace Abstract: Background: Metabolic dysfunction-associated steatotic liver disease (MASLD) affects 30-40% of US adults and is the most common chronic liver diseas

QuanBench+: A Unified Multi-Framework Benchmark for LLM-Based Quantum Code Generation

Model ReleasesDGX agent

arXiv:2604.08570v1 Announce Type: cross Abstract: Large Language Models (LLMs) are increasingly used for code generation, yet quantum code generation is still evaluated mostly within single frameworks

See you this Wednesday at the Ollama Gemma Meetup! 💎

Model ReleasesDGX agent

Ollama announced a meetup focused on Gemma, Google's open-weight language model family, scheduled for Wednesday. The event likely brought together developers and AI enthusiasts to explore running Gemm

SPASM: Stable Persona-driven Agent Simulation for Multi-turn Dialogue Generation

Model ReleasesDGX agent

arXiv:2604.09212v1 Announce Type: new Abstract: Large language models are increasingly deployed in multi-turn settings such as tutoring, support, and counseling, where reliability depends on preservin

Structured Exploration and Exploitation of Label Functions for Automated Data Annotation

ResearchDGX agent

arXiv:2604.08578v1 Announce Type: cross Abstract: High-quality labeled data is critical for training reliable machine learning and deep learning models, yet manual annotation remains costly and error-

Structured Uncertainty guided Clarification for LLM Agents

Model ReleasesDGX agent

arXiv:2511.08798v2 Announce Type: replace-cross Abstract: LLM agents with tool-calling capabilities often fail when user instructions are ambiguous or incomplete, leading to incorrect invocations and

Temporal Patch Shuffle (TPS): Leveraging Patch-Level Shuffling to Boost Generalization and Robustness in Time Series Forecasting

Local AiDGX agent

arXiv:2604.09067v1 Announce Type: new Abstract: Data augmentation is a crucial technique for improving model generalization and robustness, particularly in deep learning models where training data is

There's a speed and focus tradeoff when building a core product and just using a frontier model/harness, but this is a pretty compelling arg…

AgentsDGX agent

Harrison Chase discusses the tradeoff between speed and focus when building a core product using a frontier model versus developing more customized solutions. Using a frontier model with a standard ha

TRU: Targeted Reverse Update for Efficient Multimodal Recommendation Unlearning

Model ReleasesDGX agent

arXiv:2604.02183v2 Announce Type: replace Abstract: Multimodal recommendation systems (MRS) jointly model user-item interaction graphs and rich item content, but this tight coupling makes user data di

12 Apr 2026

Benchmark Your Local LLMs in 3 Commands

Model ReleasesDGX agent

This r/ollama post describes a streamlined method for performance-testing locally-running large language models using the Ollama framework, achievable with just three terminal commands. It likely intr

Harness, Memory, Context Fragments, & the Bitter Lesson this is a work in progress mental dump on interesting intersections between how we u…

Model ReleasesDGX agent

Harness, Memory, Context Fragments, & the Bitter Lesson this is a work in progress mental dump on interesting intersections between how we use and design a harness, implications for memory being accum

I built a VS Code extension that cuts my Claude API bill to ~$5/day

Model ReleasesDGX agent

A developer shared on r/ollama how they built a custom VS Code extension that routes Claude API requests through locally-run models via Ollama, dramatically reducing cloud API costs to approximately $

MiniMax M2.7 Advances Scalable Agentic Workflows on NVIDIA Platforms for Complex AI Applications

HardwareDGX agent

MiniMax M2.7 is an enhancement of the MiniMax M2.5 model, built as a 230B-parameter Mixture-of-Experts (MoE) model with 10B active parameters per token and a 200K context length, designed for agentic

Run it on the AI Native Cloud — serverless and dedicated infrastructure. https://www.together.ai/models/minimax-m2-7

ToolsDGX agent

MiniMax-M2 is a large-scale mixture-of-experts (MoE) language model available for inference on Together AI's platform, accessible via both serverless and dedicated infrastructure options. The model ca

11 Apr 2026

Hallucination problem

Local AiDGX agent

A Reddit thread in r/ollama where a user reports experiencing AI hallucination issues when running local language models through Ollama. The discussion likely covers symptoms such as models generating

“Maybe I should try Chat again after only using Claude for a while”. First response:

Model ReleasesDGX agent

I was unable to retrieve the specific Reddit thread content from that URL through my search. Reddit threads often require direct access to load user-generated content, and the search did not return...

Open Everything 🤝 Own your intelligence 🤝 Builder Choice We’re in the v0.1 of deploying agentic intelligence across the economy. Agents ar…

Model ReleasesDGX agent

Open Everything 🤝 Own your intelligence 🤝 Builder Choice We’re in the v0.1 of deploying agentic intelligence across the economy. Agents are data generating beasts! Experiential Memory: Each piece of g

OpenClaw + Ollama + gemma4:26b is fast in raw Ollama, but first heavy OpenClaw turns are extremely slow or hit idle timeout

Local AiDGX agent

Users running OpenClaw with `gemma4:26b` via Ollama encounter significantly slow or timed-out first turns in a session, even though the model responds quickly when queried directly through raw Olla...

10 Apr 2026

ADAG: Automatically Describing Attribution Graphs

Model ReleasesDGX agent

arXiv:2604.07615v1 Announce Type: new Abstract: In language model interpretability research, extbf{circuit tracing} aims to identify which internal features causally contributed to a particular outp

AgentOpt v0.1 Technical Report: Client-Side Optimization for LLM-Based Agent

Model ReleasesDGX agent

arXiv:2604.06296v1 Announce Type: cross Abstract: AI agents are increasingly deployed in real-world applications, including systems such as Manus, OpenClaw, and coding agents. Existing research has pr

AVGen-Bench: A Task-Driven Benchmark for Multi-Granular Evaluation of Text-to-Audio-Video Generation

Model ReleasesDGX agent

arXiv:2604.08540v1 Announce Type: cross Abstract: Text-to-Audio-Video (T2AV) generation is rapidly becoming a core interface for media creation, yet its evaluation remains fragmented. Existing benchma

Cannot search pdf document using WebUI and Ollama

Local AiDGX agent

Users in the Ollama/Open WebUI community commonly report being unable to query PDF documents via the Open WebUI interface when using a locally hosted Ollama backend, with the model failing to recog...

Chunks as Arms: Multi-Armed Bandit-Guided Sampling for Long-Context LLM Preference Optimization

Model ReleasesDGX agent

arXiv:2508.13993v2 Announce Type: replace Abstract: Long-context modeling is critical for a wide range of real-world tasks, including long-context question answering, summarization, and complex reason

ClawsBench: Evaluating Capability and Safety of LLM Productivity Agents in Simulated Workspaces

Model ReleasesDGX agent

arXiv:2604.05172v2 Announce Type: replace Abstract: Large language model (LLM) agents are increasingly deployed to automate productivity tasks (e.g., email, scheduling, document management), but evalu

Cross-Modal Emotion Transfer for Emotion Editing in Talking Face Video

TutorialsDGX agent

arXiv:2604.07786v1 Announce Type: new Abstract: Talking face generation has gained significant attention as a core application of generative models. To enhance the expressiveness and realism of synthe

Drift-Based Policy Optimization: Native One-Step Policy Learning for Online Robot Control

SafetyDGX agent

arXiv:2604.03540v2 Announce Type: replace Abstract: Although multi-step generative policies achieve strong performance in robotic manipulation by modeling multimodal action distributions, they require

DSCA: Dynamic Subspace Concept Alignment for Lifelong VLM Editing

Model ReleasesDGX agent

arXiv:2604.07965v1 Announce Type: new Abstract: Model editing aims to update knowledge to add new concepts and change relevant information without retraining. Lifelong editing is a challenging task, p

EditCaption: Human-Aligned Instruction Synthesis for Image Editing via Supervised Fine-Tuning and Direct Preference Optimization

Model ReleasesDGX agent

arXiv:2604.08213v1 Announce Type: new Abstract: High-quality training triplets (source-target image pairs with precise editing instructions) are a critical bottleneck for scaling instruction-guided im

Efficient and Effective Internal Memory Retrieval for LLM-Based Healthcare Prediction

Model ReleasesDGX agent

arXiv:2604.07659v1 Announce Type: new Abstract: Large language models (LLMs) hold significant promise for healthcare, yet their reliability in high-stakes clinical settings is often compromised by hal

Efficient PRM Training Data Synthesis via Formal Verification

ResearchDGX agent

arXiv:2505.15960v3 Announce Type: replace Abstract: Process Reward Models (PRMs) have emerged as a promising approach for improving LLM reasoning capabilities by providing process supervision over rea

Explainable AI to Improve Machine Learning Reliability for Industrial Cyber-Physical Systems

SafetyDGX agent

arXiv:2601.16074v2 Announce Type: replace Abstract: Industrial Cyber-Physical Systems (CPS) are sensitive infrastructure from both safety and economics perspectives, making their reliability criticall

FIT: A Large-Scale Dataset for Fit-Aware Virtual Try-On

Model ReleasesDGX agent

arXiv:2604.08526v1 Announce Type: new Abstract: Given a person and a garment image, virtual try-on (VTO) aims to synthesize a realistic image of the person wearing the garment, while preserving their

Flemme: A Flexible and Modular Learning Platform for Medical Images

ResearchDGX agent

arXiv:2408.09369v3 Announce Type: replace-cross Abstract: As the rapid development of computer vision and the emergence of powerful network backbones and architectures, the application of deep learnin

GenLCA: 3D Diffusion for Full-Body Avatars from In-the-Wild Videos

ApplicationsDGX agent

arXiv:2604.07273v2 Announce Type: replace Abstract: We present GenLCA, a diffusion-based generative model for generating and editing photorealistic full-body avatars from text and image inputs. The ge

HEX: Humanoid-Aligned Experts for Cross-Embodiment Whole-Body Manipulation

ApplicationsDGX agent

arXiv:2604.07993v1 Announce Type: new Abstract: Humans achieve complex manipulation through coordinated whole-body control, whereas most Vision-Language-Action (VLA) models treat robot body parts larg

Holistic Optimal Label Selection for Robust Prompt Learning under Partial Labels

Model ReleasesDGX agent

arXiv:2604.06614v1 Announce Type: cross Abstract: Prompt learning has gained significant attention as a parameter-efficient approach for adapting large pre-trained vision-language models to downstream

How SAP Concur automates expense reporting with agentic AI

Model ReleasesDGX agent

For decades, expense automation relied on a simple premise: If the machine can read the text, it can do the work. But anyone who has ever tried to scan a crumpled, smudged, or sun-bleached receipt fro

Is the ASUS ROG Flow Z13 with 128GB of Unified Memory (AMD Strix Halo) a good option to run large LLMs (70B+)?

Local AiDGX agent

The ASUS ROG Flow Z13 (2025) with AMD Ryzen AI Max+ 395 (Strix Halo) and 128GB of unified LPDDR5X memory is a capable portable option for running large LLMs locally, with ASUS officially stating it...

Lang2Act: Fine-Grained Visual Reasoning through Self-Emergent Linguistic Toolchains

ResearchDGX agent

arXiv:2602.13235v2 Announce Type: replace-cross Abstract: Visual Retrieval-Augmented Generation (VRAG) enhances Vision-Language Models (VLMs) by incorporating external visual documents to address a gi

Limits of Difficulty Scaling: Hard Samples Yield Diminishing Returns in GRPO-Tuned SLMs

SafetyDGX agent

arXiv:2604.06298v1 Announce Type: new Abstract: Recent alignment work on Large Language Models (LLMs) suggests preference optimization can improve reasoning by shifting probability mass toward better

LongSpec: Long-Context Lossless Speculative Decoding with Efficient Drafting and Verification

ResearchDGX agent

arXiv:2502.17421v4 Announce Type: replace-cross Abstract: As Large Language Models (LLMs) can now process extremely long contexts, efficient inference over these extended inputs has become increasingl

LoRA-DA: Data-Aware Initialization for Low-Rank Adaptation via Asymptotic Analysis

Model ReleasesDGX agent

arXiv:2510.24561v2 Announce Type: replace-cross Abstract: LoRA has become a widely adopted method for PEFT, and its initialization methods have attracted increasing attention. However, existing method

← Previous
1…323324325326327…1042
Next →