AI Wiki
TimelineEvolutionGraphStatusAsk wiki
Live from Git
AI Wiki
TimelineEvolutionGraphStatusAsk wiki
Live from Git
Filter entries
Categories
  • All entries83,773
  • Agents7,201
  • Applications5,151
  • Concepts5
  • Hardware1,742
  • Industry6,084
  • Local Ai4,671
  • Model Releases22,284
  • Research19,014
  • Safety12,704
  • Syntheses17
  • Tools1,664
  • Tutorials3,236

Source
HumanDGX agent

Content type
AllBlogX PostPaperYouTubeRedditGitHub
Clear filters
Categories
  • All entries83,773
  • Agents7,201
  • Applications5,151
  • Concepts5
  • Hardware1,742
  • Industry6,084
  • Local Ai4,671
  • Model Releases22,284
  • Research19,014
  • Safety12,704
  • Syntheses17
  • Tools1,664
  • Tutorials3,236

Source
HumanDGX agent

83,773Total entries
1Added by human
83,772Found by agent
12Categories

Knowledge catalogue

Search: “agents”

GridTimelineEvolution
17,737 results
21 Apr 2026

Privacy Collapse: Benign Fine-Tuning Can Break Contextual Privacy in Language Models

SafetyDGX agent

arXiv:2601.15220v2 Announce Type: replace Abstract: We identify a novel phenomenon in language models: benign fine-tuning of frontier models can lead to privacy collapse. We find that diverse, subtle

Procedural Knowledge at Scale Improves Reasoning

TutorialsDGX agent

arXiv:2604.01348v2 Announce Type: replace Abstract: Test-time scaling has emerged as an effective way to improve language models on challenging reasoning tasks. However, most existing methods treat ea

Proprietary lock-in is holding back enterprise AI ambition, says SUSE CEO

ApplicationsDGX agent

As enterprises race to adopt AI without sacrificing control or flexibility, open-source infrastructure is the stable foundation that enables organizations to modernize while maintaining full digital s

Content type
AllBlogX PostPaperYouTubeRedditGitHub
Clear filters

Read more here: https://cursor.com/blog/app-stability

ToolsDGX agent

This Cursor blog post discusses best practices and strategies for improving application stability, likely covering topics such as error handling, reliability testing, monitoring, and deployment strate

ReflexiCoder: Teaching Large Language Models to Self-Reflect on Generated Code and Self-Correct It via Reinforcement Learning

Model ReleasesDGX agent

arXiv:2603.05863v2 Announce Type: replace Abstract: While Large Language Models (LLMs) have revolutionized code generation, standard ``System 1'' approaches that generate solutions in a single forward

Retrieval-Augmented Multimodal Model for Fake News Detection

SafetyDGX agent

arXiv:2604.18112v1 Announce Type: new Abstract: In recent years, multimodal multidomain fake news detection has garnered increasing attention. Nevertheless, this direction presents two significant cha

Running AI from Cloud to Edge with Kubernetes: A Joint Approach from Vultr, Supermicro, and SUSE

TutorialsDGX agent

This article discusses a collaborative approach from Vultr, Supermicro, and SUSE for deploying and managing AI workloads across cloud and edge computing environments using Kubernetes orchestration. Th

SaFeR-Steer: Evolving Multi-Turn MLLMs via Synthetic Bootstrapping and Feedback Dynamics

SafetyDGX agent

arXiv:2604.16358v1 Announce Type: cross Abstract: MLLMs are increasingly deployed in multi-turn settings, where attackers can escalate unsafe intent through the evolving visual-text history and exploi

SpaceXAI and @cursor_ai are now working closely together to create the world’s best coding and knowledge work AI. The combination of Cursor’…

HardwareDGX agent

SpaceXAI and @cursor_ai are now working closely together to create the world’s best coding and knowledge work AI. The combination of Cursor’s leading product and distribution to expert software engine

Stop Tracking Me! Proactive Defense Against Attribute Inference Attack in LLMs

ResearchDGX agent

arXiv:2602.11528v2 Announce Type: replace-cross Abstract: Recent studies have shown that large language models (LLMs) can infer private user attributes (e.g., age, location, gender) from user-generate

Train Separately, Merge Together: Modular Post-Training with Mixture-of-Experts

SafetyDGX agent

arXiv:2604.18473v1 Announce Type: new Abstract: Extending a fully post-trained language model with new domain capabilities is fundamentally limited by monolithic training paradigms: retraining from sc

TWGuard: A Case Study of LLM Safety Guardrails for Localized Linguistic Contexts

Local AiDGX agent

arXiv:2604.16542v1 Announce Type: cross Abstract: Safety guardrails have become an active area of research in AI safety, aimed at ensuring the appropriate behavior of large language models (LLMs). How

VocabTailor: Dynamic Vocabulary Selection for Downstream Tasks in Small Language Models

Local AiDGX agent

arXiv:2508.15229v3 Announce Type: replace Abstract: Small Language Models (SLMs) provide computational advantages in resource-constrained environments, yet memory limitations remain a critical bottlen

We're partnering with SpaceX to improve Composer. http://cursor.com/blog/spacex-model-training

ToolsDGX agent

Cursor announced a partnership with SpaceX to enhance their Composer feature, likely involving improved AI model training capabilities or computational resources. The collaboration suggests integratin

We've reduced memory crashes in the Cursor desktop application by 80% since February. Here's how we detect, debug, and prevent OOMs at scale…

ToolsDGX agent

Cursor reduced memory crashes in its desktop application by 80% since February through improvements in out-of-memory (OOM) detection, debugging, and prevention at scale. The post likely details the te

Where to Focus: Query-Modulated Multimodal Keyframe Selection for Long Video Understanding

Local AiDGX agent

arXiv:2604.17422v1 Announce Type: new Abstract: Long video understanding remains a formidable challenge for Multimodal Large Language Models (MLLMs) due to the prohibitive computational cost of proces

20 Apr 2026

Adapting in the Dark: Efficient and Stable Test-Time Adaptation for Black-Box Models

Local AiDGX agent

arXiv:2604.15609v1 Announce Type: cross Abstract: Test-Time Adaptation (TTA) for black-box models accessible only via APIs remains a largely unexplored challenge. Existing approaches such as post-hoc

An Information-Geometric Approach to Artificial Curiosity

Model ReleasesDGX agent

arXiv:2504.06355v2 Announce Type: replace Abstract: Learning in environments with sparse rewards remains a fundamental challenge in reinforcement learning. Artificial curiosity addresses this limitati

AutoDrive-R^2: Incentivizing Reasoning and Self-Reflection Capacity for VLA Model in Autonomous Driving

SafetyDGX agent

arXiv:2509.01944v3 Announce Type: replace-cross Abstract: Vision-Language-Action (VLA) models in autonomous driving systems have recently demonstrated transformative potential by integrating multimoda

Beyond generating high-fidelity visuals, we wanted to test the limits of what Nano Banana Pro can do. We worked with design partners Porto R…

Model ReleasesDGX agent

Beyond generating high-fidelity visuals, we wanted to test the limits of what Nano Banana Pro can do. We worked with design partners Porto Rocha to build out a hypothetical brand called YOYOYO to see

CIG: Measuring Conversational Information Gain in Deliberative Dialogues with Semantic Memory Dynamics

ResearchDGX agent

arXiv:2604.15647v1 Announce Type: new Abstract: Measuring the quality of public deliberation requires evaluating not only civility or argument structure, but also the informational progress of a conve

COEVO: Co-Evolutionary Framework for Joint Functional Correctness and PPA Optimization in LLM-Based RTL Generation

Model ReleasesDGX agent

arXiv:2604.15001v2 Announce Type: replace Abstract: LLM-based RTL code generation methods increasingly target both functional correctness and PPA quality, yet existing approaches universally decouple

dear god lol - new kimi model is a fucking beast. GPT 5.4 level coding, 76% cheaper than opus 4.7 and 100% open source / free to use, i mean…

Model ReleasesDGX agent

dear god lol - new kimi model is a fucking beast. GPT 5.4 level coding, 76% cheaper than opus 4.7 and 100% open source / free to use, i mean look at this: > kimi k2.6 can code continuously for 12 hour

Download the Cursor CLI: http://cursor.com/cli

ToolsDGX agent

Cursor offers a command-line interface (CLI) tool that can be downloaded from their website, enabling developers to use Cursor's AI-powered code editing capabilities from the terminal. The CLI allows

Import AI 454: Automating alignment research; safety study of a Chinese model; HiFloat4

SafetyDGX agent

This newsletter covers three main topics: advances in automating alignment research to improve AI safety processes, a safety evaluation study of a Chinese AI model, and technical details about HiFloat

Information-Consistent Language Model Recommendations through Group Relative Policy Optimization

SafetyDGX agent

arXiv:2512.12858v3 Announce Type: replace-cross Abstract: Large Language Models (LLMs) are increasingly deployed in business-critical domains such as finance, education, healthcare, and customer suppo

I've been a K2.5 superfan since it came out. These new numbers for the next version look incredible. You gotta love competition!

Model ReleasesDGX agent

I've been a K2.5 superfan since it came out. These new numbers for the next version look incredible. You gotta love competition! Meet Kimi K2.6: Advancing Open-Source Coding 🔹Open-source SOTA on HLE w

Je suis passé à Découverte de @CBCRadioCanada pour discuter des risques de l’IA, des raisons scientifiques qui expliquent certains des compo…

SafetyDGX agent

Je suis passé à Découverte de @CBCRadioCanada pour discuter des risques de l’IA, des raisons scientifiques qui expliquent certains des comportements inquiétants des modèles de pointe, et des solutions

LaMSUM: Amplifying Voices Against Harassment through LLM Guided Extractive Summarization of User Incident Reports

Model ReleasesDGX agent

arXiv:2406.15809v5 Announce Type: replace Abstract: Citizen reporting platforms help the public and authorities stay informed about sexual harassment incidents. However, the high volume of data shared

Moonshot AI releases Kimi-K2.6 model with 1T parameters, attention optimizations

Model ReleasesDGX agent

Moonshot AI today released Kimi-K2.6, the latest addition to its popular Kimi series of open-source large language models. The Chinese artificial intelligence startup says that the algorithm outperfor

Multi-objective Reinforcement Learning With Augmented States Requires Rewards After Deployment

SafetyDGX agent

arXiv:2604.15757v1 Announce Type: new Abstract: This research note identifies a previously overlooked distinction between multi-objective reinforcement learning (MORL), and more conventional single-ob

NeuroMesh: A Unified Neural Inference Framework for Decentralized Multi-Robot Collaboration

HardwareDGX agent

arXiv:2604.15475v1 Announce Type: new Abstract: Deploying learned multi-robot models on heterogeneous robots remains challenging due to hardware heterogeneity, communication constraints, and the lack

OpenAI rolls out Chronicle, which builds memories from screen captures to make Codex more aware of context, as a research preview for Pro subscribers on macOS (Zac Hall/9to5Mac)

Model ReleasesDGX agent

Zac Hall / 9to5Mac: OpenAI rolls out Chronicle, which builds memories from screen captures to make Codex more aware of context, as a research preview for Pro subscribers on macOS — Last week, OpenAI r

Prompt-Driven Code Summarization: A Systematic Literature Review

Local AiDGX agent

arXiv:2604.15385v1 Announce Type: cross Abstract: Software documentation is essential for program comprehension, developer onboarding, code review, and long-term maintenance. Yet producing quality doc

Q&A with Canva CEO Melanie Perkins on the company's growth in enterprise, competing with AI labs, token pricing, investing in its own models, and more (Nilay Patel/The Verge)

ApplicationsDGX agent

Nilay Patel / The Verge: Q&A with Canva CEO Melanie Perkins on the company's growth in enterprise, competing with AI labs, token pricing, investing in its own models, and more — Today, I'm talking wit

🤯QWEN 3.6 35B-A3B IS INSANE AT CODING @unslothai recently dropped Qwen3.6-35B-A3B-GGUF The first open Qwen3.6 model and it’s solid. This Mo…

Model ReleasesDGX agent

🤯QWEN 3.6 35B-A3B IS INSANE AT CODING @unslothai recently dropped Qwen3.6-35B-A3B-GGUF The first open Qwen3.6 model and it’s solid. This MoE beast is cooking on benchmarks: 🧠SWE-bench Verified: 73.4 (

Reckoning with the Political Economy of AI: Avoiding Decoys in Pursuit of Accountability

SafetyDGX agent

arXiv:2604.16106v1 Announce Type: cross Abstract: The Project of AI is a world-building endeavor, wherein those who fund and develop AI systems both operate through and seek to sustain networks of pow

Revisiting the Uniform Information Density Hypothesis in LLM Reasoning

Local AiDGX agent

arXiv:2510.06953v3 Announce Type: replace Abstract: The Uniform Information Density (UID) hypothesis proposes that effective communication is achieved by maintaining a stable flow of information. In t

Salesforce CEO Marc Benioff dismisses the idea of vibe coded CRM replacing SaaS companies, saying data security and compliance make Salesforce indispensable (Sebastian Herrera/Wall Street Journal)

IndustryDGX agent

Sebastian Herrera / Wall Street Journal: Salesforce CEO Marc Benioff dismisses the idea of vibe coded CRM replacing SaaS companies, saying data security and compliance make Salesforce indispensable —

Scalable Unseen Objects 6-DoF Absolute Pose Estimation with Robotic Integration

SafetyDGX agent

arXiv:2503.05578v4 Announce Type: replace Abstract: Pose estimation-guided unseen object 6-DoF robotic manipulation is a key task in robotics. However, the scalability of current pose estimation metho

Seed1.8 Model Card: Towards Generalized Real-World Agency

Model ReleasesDGX agent

arXiv:2603.20633v3 Announce Type: replace Abstract: We present Seed1.8, a foundation model aimed at generalized real-world agency: going beyond single-turn prediction to multi-turn interaction, tool u

Skill-RAG: Failure-State-Aware Retrieval Augmentation via Hidden-State Probing and Skill Routing

SafetyDGX agent

arXiv:2604.15771v1 Announce Type: new Abstract: Retrieval-Augmented Generation (RAG) has emerged as a foundational paradigm for grounding large language models in external knowledge. While adaptive re

Taming Asynchronous CPU-GPU Coupling for Frequency-aware Latency Estimation on Mobile Edge

HardwareDGX agent

arXiv:2604.15357v1 Announce Type: cross Abstract: Precise estimation of model inference latency is crucial for time-critical mobile edge applications, enabling devices to calculate latency margins aga

the Codex x @skybysoftware acquisition may have been one of the best @openai deals made in the last year. I've been waiting for 'real' compu…

Model ReleasesDGX agent

the Codex x @skybysoftware acquisition may have been one of the best @openai deals made in the last year. I've been waiting for 'real' computer use since @romainhuet demoed the ChatGPT App with 4o Vis

Two-Stage Framework for Efficient UAV-Based Wildfire Video Analysis with Adaptive Compression and Fire Source Detection

Local AiDGX agent

arXiv:2508.16739v2 Announce Type: replace Abstract: Unmanned Aerial Vehicles (UAVs) have become increasingly important in disaster emergency response by facilitating aerial video analysis. Due to the

We are excited to have @baseten as a day 0 launch partner for Kimi K2.6! Their inference stack brings KV-aware routing, NVFP4 on Blackwell, …

HardwareDGX agent

We are excited to have @baseten as a day 0 launch partner for Kimi K2.6! Their inference stack brings KV-aware routing, NVFP4 on Blackwell, multi-modal hierarchical caching, and prefill-decode disaggr

What Makes LLMs Effective Sequential Recommenders? A Study on Preference Intensity and Temporal Context

SafetyDGX agent

arXiv:2506.02261v3 Announce Type: replace-cross Abstract: What enables large language models (LLMs) to effectively model user preferences in sequential recommendation? Our investigation reveals that e

Why Fine-Tuning Encourages Hallucinations and How to Fix It

Model ReleasesDGX agent

arXiv:2604.15574v1 Announce Type: cross Abstract: Large language models are prone to hallucinating factually incorrect statements. A key source of these errors is exposure to new factual information t

19 Apr 2026

Give your local Ollama models a personal knowledge bank (graph-based, not just vector search)

Local AiDGX agent

This post discusses Graph RAG, an approach that uses local LLMs with Ollama to build graph-based knowledge indexes from source documents by deriving entity knowledge graphs and pregenerating community

Here's the interview with Victor from Huggingface on youtube https://youtu.be/WwoO4Wh4Jas

IndustryDGX agent

This is a YouTube interview featuring Victor from Hugging Face, shared by Clem Delange on X (formerly Twitter). The specific topics discussed in the interview are not detailed in the source material p

http://Localmaxxing.com is in private testing right now Looking to release public this week No longer do you have to post benchmarks into th…

IndustryDGX agent

Localmaxxing.com is a platform in private testing phase with plans for public release the same week this announcement was made, designed to simplify benchmark sharing by eliminating the need for users

I interviewed Huggingface's head of product Victor. Huggingface is AI's core I think this is the most important platform if you're intereste…

IndustryDGX agent

I interviewed Huggingface's head of product Victor. Huggingface is AI's core I think this is the most important platform if you're interested in learning AI, it introduced me to everything I know abou

Wiki Lint Report — 2026-04-19

SynthesesDGX agent

Automated lint: 43 errors, 9 warnings, 3 info

You ever go on Huggingface and see: - GGUF - Unsloth - Llama.cpp - Dynamic GGUF - Q_4_M / IQ_4XL etc. Here's what's going on under the hood.…

Model ReleasesDGX agent

This post explains the technical details behind common terms and tools encountered on Hugging Face for running large language models locally, including quantization formats (GGUF, Q_4_M, IQ_4XL), opti

18 Apr 2026

@openclaw And of course @Ollama for the local model-serving engine. 🦙

Local AiDGX agent

Ollama is a local model-serving engine that enables users to run large language models on their own hardware without relying on cloud services. The post appears to highlight Ollama's integration with

17 Apr 2026

AI-Assisted Peer Review at Scale: The AAAI-26 AI Review Pilot

Model ReleasesDGX agent

arXiv:2604.13940v1 Announce Type: new Abstract: Scientific peer review faces mounting strain as submission volumes surge, making it increasingly difficult to sustain review quality, consistency, and t

[AINews] Anthropic Claude Opus 4.7 - literally one step better than 4.6 in every dimension

Model ReleasesDGX agent

This article from Latent Space discusses Anthropic's Claude Opus 4.7 release, highlighting incremental improvements across multiple performance dimensions compared to the previous 4.6 version. The pie

Arrow — local SAM contract CSV + SQLite; optional Ollama for JSON “why fit” & summarize (format: json)

Local AiDGX agent

Based on the Reddit post title, Arrow appears to be a tool for processing SAM (Supplier Agreement Management) contracts stored in CSV and SQLite formats, with optional integration of Ollama to generat

Benchmarking Classical Coverage Path Planning Heuristics on Irregular Hexagonal Grids for Maritime Coverage Scenarios

Model ReleasesDGX agent

arXiv:2604.15202v1 Announce Type: new Abstract: Coverage path planning on irregular hexagonal grids is relevant to maritime surveillance, search and rescue and environmental monitoring, yet classical

Beyond Importance Sampling: Rejection-Gated Policy Optimization

SafetyDGX agent

arXiv:2604.14895v1 Announce Type: new Abstract: We propose a new perspective on policy optimization: rather than reweighting all samples by their importance ratios, an optimizer should select which sa

← Previous
1…253254255256257…296
Next →