A Quantitative Definition of Intelligence
arXiv:2604.10873v1 Announce Type: new Abstract: We propose an operational, quantitative definition of intelligence for arbitrary physical systems. The intelligence density of a system is the ratio of
Knowledge catalogue
arXiv:2604.10873v1 Announce Type: new Abstract: We propose an operational, quantitative definition of intelligence for arbitrary physical systems. The intelligence density of a system is the ratio of
arXiv:2604.10096v1 Announce Type: new Abstract: Current embodied intelligent systems still face a substantial gap between high-level reasoning and low-level physical execution in open-world environmen
arXiv:2604.09747v1 Announce Type: cross Abstract: Large Language Model (LLM) agents have achieved rapid adoption and demonstrated remarkable capabilities across a wide range of applications. To improv
arXiv:2407.11764v2 Announce Type: replace Abstract: Existing studies have shown that Message-Passing Graph Neural Networks (MPNNs) are highly susceptible to adversarial attacks. In contrast, despite t
arXiv:2604.09633v1 Announce Type: cross Abstract: This work examines how AI, especially agentic systems, is being adopted in engineering and manufacturing workflows, what value it provides today, and
arXiv:2604.10383v1 Announce Type: new Abstract: Existing multi-agent video generation systems use LLM agents to orchestrate neural video generators, producing visually impressive but semantically unre
arXiv:2604.09576v1 Announce Type: new Abstract: Deploying continual object detection on microcontrollers (MCUs) with under 100KB memory requires efficient feature compression that can adapt to evolvin
This Reddit thread from r/StableDiffusion discusses the community's search for AI tools capable of analyzing an existing video and automatically generating a descriptive text prompt from it — essentia
Face and head swapping capabilities are available for LTX 2.3, primarily through external LoRA models and ComfyUI workflows rather than built-in native features. A dedicated face replacement workflow
arXiv:2510.17934v2 Announce Type: replace-cross Abstract: Retrieval-augmented generation (RAG) has shown some success in augmenting large language models (LLMs) with external knowledge. However, as a
arXiv:2604.10598v1 Announce Type: new Abstract: Human-in-the-loop (HITL) UAV operation is essential in complex and safety-critical aerial surveying environments, where human operators provide navigati
Build b8787 is a tagged release of llama.cpp, an open-source C/C++ library for running large language model (LLM) inference locally or in the cloud with minimal setup. As with all llama.cpp builds, it
Build b8789 is an incremental release of llama.cpp, the open-source C/C++ library for running large language model inference locally and in the cloud. Like other builds in the project's continuous rel
Build b8793 is a tagged release of the llama.cpp project, an open-source C/C++ framework for running large language model (LLM) inference locally or in the cloud with minimal setup. As part of llama.c
Build b8794 is an incremental release of llama.cpp, the open-source C/C++ library for running large language model (LLM) inference locally. Like other builds in the project's continuous release series
arXiv:2604.09671v1 Announce Type: new Abstract: We propose a stronger formulation of RL on top of RWKV-style recurrent sequence models, in which the fixed-size recurrent state is explicitly interprete
This r/StableDiffusion post documents a community experiment exploring techniques for generating images of two identical-looking subjects ('twins') using Stable Diffusion without relying on a LoRA mod
arXiv:2604.10507v1 Announce Type: new Abstract: Psychological client simulators have emerged as a scalable solution for training and evaluating counselor trainees and psychological LLMs. Yet existing
arXiv:2604.02369v3 Announce Type: replace-cross Abstract: Agent communication protocols are becoming critical infrastructure for large language model (LLM) systems that must use tools, coordinate with
arXiv:2604.09549v1 Announce Type: cross Abstract: Recommender systems are central to online services, enabling users to navigate through massive amounts of content across various domains. However, the
arXiv:2604.09927v1 Announce Type: new Abstract: Robust license plate recognition in unconstrained environments remains a significant challenge, particularly in underrepresented regions with limited da
arXiv:2604.10900v1 Announce Type: new Abstract: In large language models performing long-form reasoning, the KV cache grows rapidly with decode length, creating bottlenecks in memory and inference sta
arXiv:2604.10502v1 Announce Type: new Abstract: Content moderation in online platforms faces persistent challenges due to the evolving complexity of user-generated content and the limitations of tradi
arXiv:2604.10114v1 Announce Type: cross Abstract: The generation of high-fidelity synthetic data is a cornerstone of modern machine learning, yet Large Language Models (LLMs) frequently suffer from ha
arXiv:2604.09812v1 Announce Type: new Abstract: Recurrent claims present a major challenge for automated fact-checking systems designed to combat misinformation, especially in multilingual settings. W
arXiv:2604.11784v1 Announce Type: cross Abstract: GUI agents drive applications through their visual interfaces instead of programmatic APIs, interacting with arbitrary software via taps, swipes, and
arXiv:2601.07224v2 Announce Type: replace Abstract: While Hybrid Supervised Fine-Tuning (SFT) followed by Reinforcement Learning (RL) has become the standard paradigm for training LLM agents, effectiv
arXiv:2604.11359v1 Announce Type: new Abstract: Accurate interpretation of electrocardiogram (ECG) remains challenging due to the scarcity of labeled data and the high cost of expert annotation. Self-
arXiv:2604.11165v1 Announce Type: cross Abstract: Clinical decision-making often involves selecting tests that are costly, invasive, or time-consuming, motivating individualized, sequential strategies
arXiv:2604.10503v1 Announce Type: cross Abstract: Modern audio systems universally employ mel-scale representations derived from 1940s Western psychoacoustic studies, potentially encoding cultural bia
arXiv:2604.10918v1 Announce Type: new Abstract: Tables contain rich structured information, yet when stored as images their contents remain 'locked' within pixels. Converting table images into LaTeX c
arXiv:2604.09728v1 Announce Type: new Abstract: Infrared thermography (IRT) is a widely used non-destructive testing technique for detecting structural features such as subsurface defects. However, mo
arXiv:2407.11077v3 Announce Type: replace-cross Abstract: The symmetry of dynamical systems can be exploited for state-transition prediction and to facilitate control policy optimization. This paper l
arXiv:2601.11428v4 Announce Type: replace Abstract: Neural PDE solvers have shown strong performance on standard benchmarks, but their robustness under deployment-relevant distribution shifts remains
arXiv:2604.10546v1 Announce Type: new Abstract: The rapid growth of visual data under stringent storage and bandwidth constraints makes extremely low-bitrate image compression increasingly important.
arXiv:2508.15452v3 Announce Type: replace-cross Abstract: Numerous deep learning-based solutions have been developed for the automatic recognition of breast cancer using mammography images. However, t
arXiv:2604.09599v1 Announce Type: cross Abstract: High-performance computing systems are complex machines whose behaviour is governed by the correct functioning of its many subsystems. Among these, th
arXiv:2604.10459v1 Announce Type: new Abstract: The exponential growth of user-generated movie reviews on digital platforms has made accurate text sentiment classification a cornerstone task in natura
arXiv:2604.09815v1 Announce Type: new Abstract: Computer-use agents that combine GUI interaction with structured API calls via the Model Context Protocol (MCP) show promise for automating software tas
arXiv:2604.09742v1 Announce Type: cross Abstract: Rotary Position Embedding (RoPE) has become a core component of modern Transformer architectures across language, vision, and 3D domains. However, exi
arXiv:2601.19019v2 Announce Type: replace-cross Abstract: Neural population activity in sensory cortex is organized on low-dimensional manifolds, but why such manifolds arise and what determines their
arXiv:2604.11422v1 Announce Type: cross Abstract: The ``differentiability gap'' presents a primary bottleneck in Earth system deep learning: since models cannot be trained directly on non-differentiab
arXiv:2604.10466v1 Announce Type: new Abstract: Visual feedback is critical for motor skill acquisition in sports and rehabilitation, and psychological studies show that observing near-perfect version
arXiv:2604.09622v1 Announce Type: cross Abstract: The rapid adoption of generative artificial intelligence (AI) in educational assessment has created new opportunities for scalable item creation, pers
arXiv:2603.23964v2 Announce Type: replace Abstract: The remarkable progress of reinforcement learning (RL) is intrinsically tied to the environments used to train and evaluate artificial agents. Movin
arXiv:2604.11376v1 Announce Type: cross Abstract: Removing patient-specific information from medical images is crucial to enable sharing and open science without compromising patient identities. Howev
arXiv:2604.11585v1 Announce Type: new Abstract: Multimodal perception systems for robotics and embodied AI often assume reliable RGB-D sensing, but in practice, depth is frequently missing, noisy, or
arXiv:2604.11600v1 Announce Type: new Abstract: Multimodal Large Language Models (MLLMs) have achieved remarkable progress but continue to struggle with geometric reasoning, primarily due to the perce
arXiv:2604.09993v1 Announce Type: new Abstract: Contact-implicit trajectory optimization (CITO) enables the automatic discovery of contact sequences, but most methods rely on fine time discretization
arXiv:2604.09711v1 Announce Type: cross Abstract: Multimodal fake news detection (MFND) aims to verify news credibility by jointly exploiting textual and visual evidence. However, real-world news diss
This Reddit post addresses a common ChatGPT image generation behavior where the model automatically carries over the seed from a previously generated image, causing new generations to visually resembl
This Reddit thread from r/ollama discusses user experiences comparing Ollama Cloud Pro to the free tier. Ollama Cloud is available through subscription tiers — Free, Pro at 20/month, and Max at 100/mo
A community-shared open source CLI agent project posted to r/ollama, designed specifically for use with 8K token context windows in local Ollama-based LLM setups. Version 0.3 focuses on improving Olla
IC-LoRA (In-Context LoRA) enables conditioning video generation on reference video frames at inference time, allowing fine-grained video-to-video control on top of a text-to-video base model. Unlike t
arXiv:2604.09702v1 Announce Type: cross Abstract: Precise segmentation of objects with highly similar shapes remains a challenging problem in dense prediction, especially in scenarios with ambiguous b
arXiv:2511.11938v2 Announce Type: replace-cross Abstract: Precise neutrino energy reconstruction is essential for next-generation long-baseline oscillation experiments, yet current methods remain limi
arXiv:2604.10703v1 Announce Type: new Abstract: Transformer architectures are designed by trial and error: the number of attention heads, the depth, and the head size are fixed before training begins,
arXiv:2503.23178v2 Announce Type: replace Abstract: Conflicts between humans and bears on the Tibetan Plateau present substantial threats to local communities and hinder wildlife preservation initiati
arXiv:2509.26306v4 Announce Type: replace Abstract: Existing multi-agent learning approaches have developed interactive training environments to explicitly promote collaboration among multiple Large L
arXiv:2604.10783v1 Announce Type: new Abstract: Designing reward functions remains a central challenge in reinforcement learning (RL) for healthcare, where outcomes are sparse, delayed, and difficult