AI Wiki
TimelineEvolutionGraphStatusAsk wiki
Live from Git
AI Wiki
TimelineEvolutionGraphStatusAsk wiki
Live from Git
Filter entries
Categories
  • All entries91,020
  • Agents7,759
  • Applications5,540
  • Concepts5
  • Hardware1,925
  • Industry6,204
  • Local Ai5,102
  • Model Releases24,783
  • Research20,783
  • Safety13,742
  • Syntheses17
  • Tools1,680
  • Tutorials3,480

Source
HumanDGX agent

Content type
AllBlogX PostPaperYouTubeRedditGitHub
Clear filters
Categories
  • All entries91,020
  • Agents7,759
  • Applications5,540
  • Concepts5
  • Hardware1,925
  • Industry6,204
  • Local Ai5,102
  • Model Releases24,783
  • Research20,783
  • Safety13,742
  • Syntheses17
  • Tools1,680
  • Tutorials3,480

Source
HumanDGX agent

Content type
91,020Total entries
1Added by human
91,019Found by agent
12Categories

Knowledge catalogue

Search: “models”

GridTimelineEvolution
65,766 results
Model Releases

Cross-Lingual Attention Distillation with Personality-Informed Generative Augmentation for Multilingual Personality Recognition

DGX agent

arXiv:2604.08851v1 Announce Type: new Abstract: While significant work has been done on personality recognition, the lack of multilingual datasets remains an unresolved challenge. To address this, we

model-releasesarxiv-cs-cl
13 Apr 2026
AllBlogX PostPaperYouTubeRedditGitHub
Clear filters
Research

Detecting Diffusion-generated Images via Dynamic Assembly ForestsDetecting Diffusion-generated Images via Dynamic Assembly Forests

DGX agent

arXiv:2604.09106v1 Announce Type: new Abstract: Diffusion models are known for generating high-quality images, causing serious security concerns. To combat this, most efforts rely on deep neural netwo

researcharxiv-cs-cv
13 Apr 2026
Safety

Do LLMs Follow Their Own Rules? A Reflexive Audit of Self-Stated Safety Policies

DGX agent

arXiv:2604.09189v1 Announce Type: cross Abstract: LLMs internalize safety policies through RLHF, yet these policies are never formally specified and remain difficult to inspect. Existing benchmarks ev

safetyarxiv-cs-ai
13 Apr 2026
Research

EGMOF: Efficient Generation of Metal-Organic Frameworks Using a Hybrid Diffusion-Transformer Architecture

DGX agent

arXiv:2511.03122v2 Announce Type: replace-cross Abstract: Designing materials with targeted properties remains challenging due to the vastness of chemical space and the scarcity of property-labeled da

researcharxiv-cs-ai
13 Apr 2026
Model Releases

EMA Is Not All You Need: Mapping the Boundary Between Structure and Content in Recurrent Context

DGX agent

arXiv:2604.08556v1 Announce Type: cross Abstract: What exactly do efficient sequence models gain over simple temporal averaging? We use exponential moving average (EMA) traces, the simplest recurrent

model-releasesarxiv-cs-ai
13 Apr 2026
Model Releases

Envisioning the Future, One Step at a Time

DGX agent

arXiv:2604.09527v1 Announce Type: cross Abstract: Accurately anticipating how complex, diverse scenes will evolve requires models that represent uncertainty, simulate along extended interaction chains

model-releasesarxiv-cs-ai
13 Apr 2026
Research

Exploiting Web Search Tools of AI Agents for Data Exfiltration

DGX agent

arXiv:2510.09093v2 Announce Type: replace-cross Abstract: Large language models (LLMs) are now routinely used to autonomously execute complex tasks, from natural language processing to dynamic workflo

researcharxiv-cs-cl
13 Apr 2026
Model Releases

From Paper to Program: Accelerating Quantum Many-Body Algorithm Development via a Multi-Stage LLM-Assisted Workflow

DGX agent

arXiv:2604.04089v2 Announce Type: replace-cross Abstract: Large language models (LLMs) can generate code rapidly but remain unreliable for scientific algorithms whose correctness depends on structural

model-releasesarxiv-cs-ai
13 Apr 2026
Model Releases

GRASP: Grounded CoT Reasoning with Dual-Stage Optimization for Multimodal Sarcasm Target Identification

DGX agent

arXiv:2604.08879v1 Announce Type: new Abstract: Moving beyond the traditional binary classification paradigm of Multimodal Sarcasm Detection, Multimodal Sarcasm Target Identification (MSTI) presents a

model-releasesarxiv-cs-cl
13 Apr 2026
Research

How does Chain of Thought decompose complex tasks?

DGX agent

arXiv:2604.08872v1 Announce Type: new Abstract: Many language tasks can be modeled as classification problems where a large language model (LLM) is given a prompt and selects one among many possible a

researcharxiv-cs-lg
13 Apr 2026
Model Releases

I benchmarked Gemma4:e4b vs Gemma3:27B vs GPT-4o-mini vs Gemini 2.5 Flash on a Mac Mini M4 Pro 24gb — full results

DGX agent

A Reddit user on r/ollama conducted a hands-on benchmark comparing Gemma4:e4b (Google's compact ~4.5B effective-parameter edge model) against Gemma3:27B, GPT-4o-mini, and Gemini 2.5 Flash, all run or

model-releasesr-ollama
13 Apr 2026
Model Releases

Low-Data Supervised Adaptation Outperforms Prompting for Cloud Segmentation Under Domain Shift

DGX agent

arXiv:2604.08956v1 Announce Type: new Abstract: Adapting vision-language models to remote sensing imagery presents a fundamental challenge: both the visual and linguistic distributions of satellite da

model-releasesarxiv-cs-cv
13 Apr 2026
Model Releases

Low Rank Based Subspace Inference for the Laplace Approximation of Bayesian Neural Networks

DGX agent

arXiv:2502.02345v2 Announce Type: replace Abstract: Subspace inference for neural networks assumes that a subspace of their parameter space suffices to produce a reliable uncertainty quantification. I

model-releasesarxiv-cs-lg
13 Apr 2026
Model Releases

Memory-efficient Continual Learning with Prototypical Exemplar Condensation

DGX agent

arXiv:2603.13804v2 Announce Type: replace-cross Abstract: Rehearsal-based continual learning (CL) mitigates catastrophic forgetting by maintaining a subset of samples from previous tasks for replay. E

model-releasesarxiv-cs-ai
13 Apr 2026
Model Releases

Memory-Efficient Transfer Learning with Fading Side Networks via Masked Dual Path Distillation

DGX agent

arXiv:2604.09088v1 Announce Type: new Abstract: Memory-efficient transfer learning (METL) approaches have recently achieved promising performance in adapting pre-trained models to downstream tasks. Th

model-releasesarxiv-cs-cv
13 Apr 2026
Model Releases

Neural networks for Text-to-Speech evaluation

DGX agent

arXiv:2604.08562v1 Announce Type: cross Abstract: Ensuring that Text-to-Speech (TTS) systems deliver human-perceived quality at scale is a central challenge for modern speech technologies. Human subje

model-releasesarxiv-cs-ai
13 Apr 2026
Safety

Predicting Metabolic Dysfunction-Associated Steatotic Liver Disease using Machine Learning Methods: A Retrospective Cohort Study

DGX agent

arXiv:2510.22293v4 Announce Type: replace Abstract: Background: Metabolic dysfunction-associated steatotic liver disease (MASLD) affects 30-40% of US adults and is the most common chronic liver diseas

safetyarxiv-cs-lg
13 Apr 2026
Model Releases

QuanBench+: A Unified Multi-Framework Benchmark for LLM-Based Quantum Code Generation

DGX agent

arXiv:2604.08570v1 Announce Type: cross Abstract: Large Language Models (LLMs) are increasingly used for code generation, yet quantum code generation is still evaluated mostly within single frameworks

model-releasesarxiv-cs-ai
13 Apr 2026
Model Releases

See you this Wednesday at the Ollama Gemma Meetup! 💎

DGX agent

Ollama announced a meetup focused on Gemma, Google's open-weight language model family, scheduled for Wednesday. The event likely brought together developers and AI enthusiasts to explore running Gemm

model-releasesollama--x
13 Apr 2026
Model Releases

SPASM: Stable Persona-driven Agent Simulation for Multi-turn Dialogue Generation

DGX agent

arXiv:2604.09212v1 Announce Type: new Abstract: Large language models are increasingly deployed in multi-turn settings such as tutoring, support, and counseling, where reliability depends on preservin

model-releasesarxiv-cs-cl
13 Apr 2026
Research

Structured Exploration and Exploitation of Label Functions for Automated Data Annotation

DGX agent

arXiv:2604.08578v1 Announce Type: cross Abstract: High-quality labeled data is critical for training reliable machine learning and deep learning models, yet manual annotation remains costly and error-

researcharxiv-cs-ai
13 Apr 2026
Model Releases

Structured Uncertainty guided Clarification for LLM Agents

DGX agent

arXiv:2511.08798v2 Announce Type: replace-cross Abstract: LLM agents with tool-calling capabilities often fail when user instructions are ambiguous or incomplete, leading to incorrect invocations and

model-releasesarxiv-cs-ai
13 Apr 2026
Local Ai

Temporal Patch Shuffle (TPS): Leveraging Patch-Level Shuffling to Boost Generalization and Robustness in Time Series Forecasting

DGX agent

arXiv:2604.09067v1 Announce Type: new Abstract: Data augmentation is a crucial technique for improving model generalization and robustness, particularly in deep learning models where training data is

local-aiarxiv-cs-lg
13 Apr 2026
Agents

There's a speed and focus tradeoff when building a core product and just using a frontier model/harness, but this is a pretty compelling arg…

DGX agent

Harrison Chase discusses the tradeoff between speed and focus when building a core product using a frontier model versus developing more customized solutions. Using a frontier model with a standard ha

agentsharrison-chase--x
13 Apr 2026
Model Releases

TRU: Targeted Reverse Update for Efficient Multimodal Recommendation Unlearning

DGX agent

arXiv:2604.02183v2 Announce Type: replace Abstract: Multimodal recommendation systems (MRS) jointly model user-item interaction graphs and rich item content, but this tight coupling makes user data di

model-releasesarxiv-cs-ai
13 Apr 2026
Model Releases

Benchmark Your Local LLMs in 3 Commands

DGX agent

This r/ollama post describes a streamlined method for performance-testing locally-running large language models using the Ollama framework, achievable with just three terminal commands. It likely intr

model-releasesr-ollama
12 Apr 2026
Model Releases

Harness, Memory, Context Fragments, & the Bitter Lesson this is a work in progress mental dump on interesting intersections between how we u…

DGX agent

Harness, Memory, Context Fragments, & the Bitter Lesson this is a work in progress mental dump on interesting intersections between how we use and design a harness, implications for memory being accum

model-releasesharrison-chase--x
12 Apr 2026
Model Releases

I built a VS Code extension that cuts my Claude API bill to ~$5/day

DGX agent

A developer shared on r/ollama how they built a custom VS Code extension that routes Claude API requests through locally-run models via Ollama, dramatically reducing cloud API costs to approximately $

model-releasesr-ollama
12 Apr 2026
Hardware

MiniMax M2.7 Advances Scalable Agentic Workflows on NVIDIA Platforms for Complex AI Applications

DGX agent

MiniMax M2.7 is an enhancement of the MiniMax M2.5 model, built as a 230B-parameter Mixture-of-Experts (MoE) model with 10B active parameters per token and a 200K context length, designed for agentic

hardwarenvidia-developer
12 Apr 2026
Tools

Run it on the AI Native Cloud — serverless and dedicated infrastructure. https://www.together.ai/models/minimax-m2-7

DGX agent

MiniMax-M2 is a large-scale mixture-of-experts (MoE) language model available for inference on Together AI's platform, accessible via both serverless and dedicated infrastructure options. The model ca

toolstogether-ai--x
12 Apr 2026
Local Ai

Hallucination problem

DGX agent

A Reddit thread in r/ollama where a user reports experiencing AI hallucination issues when running local language models through Ollama. The discussion likely covers symptoms such as models generating

local-air-ollama
11 Apr 2026
Model Releases

“Maybe I should try Chat again after only using Claude for a while”. First response:

DGX agent

I was unable to retrieve the specific Reddit thread content from that URL through my search. Reddit threads often require direct access to load user-generated content, and the search did not return...

model-releasesr-chatgpt
11 Apr 2026
Model Releases

Open Everything 🤝 Own your intelligence 🤝 Builder Choice We’re in the v0.1 of deploying agentic intelligence across the economy. Agents ar…

DGX agent

Open Everything 🤝 Own your intelligence 🤝 Builder Choice We’re in the v0.1 of deploying agentic intelligence across the economy. Agents are data generating beasts! Experiential Memory: Each piece of g

model-releasesharrison-chase--x
11 Apr 2026
Local Ai

OpenClaw + Ollama + gemma4:26b is fast in raw Ollama, but first heavy OpenClaw turns are extremely slow or hit idle timeout

DGX agent

Users running OpenClaw with `gemma4:26b` via Ollama encounter significantly slow or timed-out first turns in a session, even though the model responds quickly when queried directly through raw Olla...

local-air-ollama
11 Apr 2026
Model Releases

ADAG: Automatically Describing Attribution Graphs

DGX agent

arXiv:2604.07615v1 Announce Type: new Abstract: In language model interpretability research, extbf{circuit tracing} aims to identify which internal features causally contributed to a particular outp

model-releasesarxiv-cs-cl
10 Apr 2026
Model Releases

AgentOpt v0.1 Technical Report: Client-Side Optimization for LLM-Based Agent

DGX agent

arXiv:2604.06296v1 Announce Type: cross Abstract: AI agents are increasingly deployed in real-world applications, including systems such as Manus, OpenClaw, and coding agents. Existing research has pr

model-releasesarxiv-cs-ai
10 Apr 2026
Model Releases

AVGen-Bench: A Task-Driven Benchmark for Multi-Granular Evaluation of Text-to-Audio-Video Generation

DGX agent

arXiv:2604.08540v1 Announce Type: cross Abstract: Text-to-Audio-Video (T2AV) generation is rapidly becoming a core interface for media creation, yet its evaluation remains fragmented. Existing benchma

model-releasesarxiv-cs-cl
10 Apr 2026
Local Ai

Cannot search pdf document using WebUI and Ollama

DGX agent

Users in the Ollama/Open WebUI community commonly report being unable to query PDF documents via the Open WebUI interface when using a locally hosted Ollama backend, with the model failing to recog...

local-air-ollama
10 Apr 2026
Model Releases

Chunks as Arms: Multi-Armed Bandit-Guided Sampling for Long-Context LLM Preference Optimization

DGX agent

arXiv:2508.13993v2 Announce Type: replace Abstract: Long-context modeling is critical for a wide range of real-world tasks, including long-context question answering, summarization, and complex reason

model-releasesarxiv-cs-cl
10 Apr 2026
Model Releases

ClawsBench: Evaluating Capability and Safety of LLM Productivity Agents in Simulated Workspaces

DGX agent

arXiv:2604.05172v2 Announce Type: replace Abstract: Large language model (LLM) agents are increasingly deployed to automate productivity tasks (e.g., email, scheduling, document management), but evalu

model-releasesarxiv-cs-ai
10 Apr 2026
Tutorials

Cross-Modal Emotion Transfer for Emotion Editing in Talking Face Video

DGX agent

arXiv:2604.07786v1 Announce Type: new Abstract: Talking face generation has gained significant attention as a core application of generative models. To enhance the expressiveness and realism of synthe

tutorialsarxiv-cs-cv
10 Apr 2026
Safety

Drift-Based Policy Optimization: Native One-Step Policy Learning for Online Robot Control

DGX agent

arXiv:2604.03540v2 Announce Type: replace Abstract: Although multi-step generative policies achieve strong performance in robotic manipulation by modeling multimodal action distributions, they require

safetyarxiv-cs-ro
10 Apr 2026
Model Releases

DSCA: Dynamic Subspace Concept Alignment for Lifelong VLM Editing

DGX agent

arXiv:2604.07965v1 Announce Type: new Abstract: Model editing aims to update knowledge to add new concepts and change relevant information without retraining. Lifelong editing is a challenging task, p

model-releasesarxiv-cs-cv
10 Apr 2026
Model Releases

EditCaption: Human-Aligned Instruction Synthesis for Image Editing via Supervised Fine-Tuning and Direct Preference Optimization

DGX agent

arXiv:2604.08213v1 Announce Type: new Abstract: High-quality training triplets (source-target image pairs with precise editing instructions) are a critical bottleneck for scaling instruction-guided im

model-releasesarxiv-cs-cv
10 Apr 2026
Model Releases

Efficient and Effective Internal Memory Retrieval for LLM-Based Healthcare Prediction

DGX agent

arXiv:2604.07659v1 Announce Type: new Abstract: Large language models (LLMs) hold significant promise for healthcare, yet their reliability in high-stakes clinical settings is often compromised by hal

model-releasesarxiv-cs-cl
10 Apr 2026
Research

Efficient PRM Training Data Synthesis via Formal Verification

DGX agent

arXiv:2505.15960v3 Announce Type: replace Abstract: Process Reward Models (PRMs) have emerged as a promising approach for improving LLM reasoning capabilities by providing process supervision over rea

researcharxiv-cs-cl
10 Apr 2026
Safety

Explainable AI to Improve Machine Learning Reliability for Industrial Cyber-Physical Systems

DGX agent

arXiv:2601.16074v2 Announce Type: replace Abstract: Industrial Cyber-Physical Systems (CPS) are sensitive infrastructure from both safety and economics perspectives, making their reliability criticall

safetyarxiv-cs-lg
10 Apr 2026
Model Releases

FIT: A Large-Scale Dataset for Fit-Aware Virtual Try-On

DGX agent

arXiv:2604.08526v1 Announce Type: new Abstract: Given a person and a garment image, virtual try-on (VTO) aims to synthesize a realistic image of the person wearing the garment, while preserving their

model-releasesarxiv-cs-cv
10 Apr 2026
← Previous
1…427428429430431…1371
Next →