Papers nuevos de Computación
En los últimos 3 días aparecieron 171 papers nuevos de Computación. Pipette los leyó todos y estos son los 60 con más interés, avance real y afirmaciones prudentes. Cada uno muestra la oración de su resumen que dice el resultado principal, tal como la escribieron sus autores.
Lo mejor de los últimos 3 días
Residential Electricity Consumption Dataset for Sri Lanka (RECON-SL)
As the first dataset of its kind in Sri Lanka, and among the few worldwide to link smart meter records with longitudinal survey data at this scale, it provides a unique resource for energy, machine learning and policy research, with known gaps in smart meter coverage documented for users.
PreprintUso en el mundo realFácil de leerA Data-Driven Analysis of Infostealer Malware Victims
To close this gap, we build a privacy-preserving pipeline that turns illicitly sourced infostealer logs into a reproducible research artifact, minimizing sensitive data while preserving measurement utility, and use it to construct a dataset of 170,298 victims from logs of multiple infostealer families.
Preprint con versión publicadaDice ser un gran avanceUso en el mundo realCodeGraph: Open-Taxonomy Knowledge Graph for Source Code with Wikidata Grounding
We applied our pipeline to the 167 million files of the Stack-Edu corpus, creating the first known large-scale open-taxonomy knowledge graph for source code.
Preprint con versión publicadaDice ser un gran avancePersistent Billable State: Denial-of-Wallet Attacks and Defenses in Tool-Calling LLM Agents
These results establish persistent billable state as a first-class security object and pre-reingestion as its host-owned control point.
PreprintDice ser un gran avanceUso en el mundo realPoster: FedWM-Guard: Thwarting Imagination Poisoning in Federated World Model-based Autonomous Driving
We present FedWM-Guard, to the best of our knowledge, the first defense to characterize planner-facing rollouts in federated WM-AD, screen authenticated updates in hidden-canary scenarios, audit predicted futures against later observations, and invoke a WM-independent safety shield when persistent inconsistency is detected.
PreprintDice ser un gran avanceUso en el mundo realPrefilling the Reasoning Channel: Output-Prefix Attacks on Reasoning LLMs
We present the first systematic, controlled study that isolates the scratchpad reasoning channel as an output-prefix attack vector, and the first to compare reasoning-only, output-prefix-only and reasoning-plus-output-prefix attacks across both exposed- and hidden-reasoning models.
PreprintUso en el mundo realCuACD: A Fully GPU-Resident Approximate Convex Decomposition
Building on this template, we present CuACD (CUDA ACD), the first fully GPU-resident ACD system, together with a suite of reusable GPU components, released as open-source standalone CUDA modules that drop into any search-based ACD pipeline.
Preprint con versión publicadaDice ser un gran avanceUso en el mundo realCodetta: High-Capacity, Keyless, and Undetectable Multi-Agent Collusion
We make the threat of undetectable agent collusion concrete with Codetta, a high-capacity steganographic protocol for independently deployed agents in realistic asymmetric settings.
PreprintDice ser un gran avanceUso en el mundo realAgent Approval Laundering: Transitive Effects Beyond the Approved Invocation
We present the first systematic security analysis of this record-to-closure relation in agent systems.
PreprintDice ser un gran avanceUso en el mundo realVQ-LIC: Shared Vector-Quantized Learned Image Compression on a Resource-Constrained FPGA
We present VQ-LIC, an asymmetric edge-cloud codec in which a compact INT8 depthwise (DW)-pointwise (PW) analysis transform and multi-codebook vector quantization (VQ) run at the edge on a reusable DW/PW engine pair, while reconstruction is handled by a larger cloud decoder.
PreprintDice ser un gran avanceUso en el mundo realTrusted Model Environment for Private Semantic Computations
We introduce trusted model environments (TME), the first such primitive that executes generative models inside trusted execution environments (TEEs) while controlling output leakage.
PreprintAfirmaciones fuertes, leer con cuidadoDice ser un gran avanceOn the Effectiveness of Kernel-Level Evidence for Agent Security
Across four distinct detector families, we find that kernel evidence is discriminative on its own and that composing it with application-layer evidence generally outperforms either single-layer view, revealing complementary signals that single-layer analyses can miss.
PreprintUso en el mundo realA Procedure for Classifying Attachments and Affective Social Bonds in Human-Robot Dyads
By replacing the divergent operationalisations with a unified, criterion-based classification, this paper gives HRI practitioners a standardised basis for evaluating, classifying, and comparing human-robot relationships, and sets out the experimental rigour that each classification demands.
PreprintXtrace: High-Fidelity GPU Intra-Kernel Tracing via Binary-Level Instruction Splicing
Xtrace is the first GPU kernel tracing system with near-zero compile-time interference and minimized runtime overhead.
PreprintAfirmaciones fuertes, leer con cuidadoDice ser un gran avanceUso en el mundo real"What I See is What I Hear": Deepfake Detection Across Diverse Hearing Abilities
Our work systematically characterizes how deepfakes affect DHH populations, highlighting the asymmetric risks audiovisual manipulations may pose to groups with different hearing abilities and the need for accessible, tailored defenses that support all users.
PreprintUso en el mundo realBeyond Driving: Envisioning Activities in Future Autonomous Vehicles through Experience-Centered Design
Our findings lead us to conceptualize NDRAs not as isolated instances of "travel time use," but as dynamic sequences of interrelated activities deeply shaped by pre- and post-journey contexts.
PreprintA New Gap Sequence for Shellsort: RL-Driven Algorithm Discovery Beyond
Thus one exact sequence connects self-supervised discovery, large-scale practical performance, and a substantial step below the classical bound for sparse practical Shellsort sequences.
PreprintAfirmaciones fuertes, leer con cuidadoDice ser un gran avanceThe Tokens Remember: When Tokenization Bypasses Knowledge Editing and Unlearning
We introduce Toketive, a simple yet powerful reference-free attack that exploits the tokenization-based side channel to (i) detect modified knowledge and (ii) reconstruct the corresponding pre-edit response.
PreprintUso en el mundo realA General Composition Theorem for Approximate Degree
We resolve this question for all total Boolean functions by proving the matching lower bound.
PreprintDice ser un gran avancePrivacy Leakage Through AI-mediated Analysis of Smartphone Data
Through an IRB-approved user study, 465 participants deployed Priva-See on their phones; Priva-See made privacy-invasive inferences despite having access to only a subset of a user's data.
PreprintUso en el mundo realFácil de leerAssembling, Breaking, and Refusing the Mask: Agency in AI-Mediated Self-Presentation in Livestreaming
We conceptualize masking as a sociotechnical assemblage in which agency lies in preserving, disrupting, and reconfiguring relations rather than controlling a single interface.
PreprintCan Labor Markets Function in the Age of AI? The Evaluation Bottleneck in Hiring
Our results show how AI can shift the central friction in hiring from submitting applications to obtaining credible evaluation, creating entry barriers for high-fit workers without prior experience.
PreprintPractical and Space-Efficient LZ77 and LZ Pre-Compression via String Synchronizing Sets
We replace both, fine-tune every remaining stage, and obtain the first practical implementation, which runs in space close to the text rather than to the suffix array.
PreprintCódigo disponibleSharp Lovasz-Theta Bounds on Random Graphs
In this work, we resolve this question by proving that the \Lovasz-Theta function of is with high probability, determining its asymptotic value up to vanishing relative error.
PreprintDice ser un gran avanceWorking with Agentic `Teammates': When a New Organizational Actor Collides with the Human Ecosystem of Work
Our findings reveal the boundaries of the human-agent workplace are actively in flux, triggering breakdowns and negotiations across: 1) tacit rules of collaborative human workflows, 2) the relational boundaries of this new non-human actor, and 3) the redistribution of trust and human agency.
PreprintDon't Read the Log: Execution Traces Contaminate Verifiers in Video-Generation Agents
On a benchmark of 109 generated two-event clips with manual labels, in which the requested event is either visibly completed or visibly missing, a trace that reports a successful tool call makes three open-weight Qwen-VL judges (7B, 8B, 32B) accept -- of the failures, up from -- without text, and a contradicting trace makes them reject up to of correct clips; an instruction to "use only the frames" does not remove the effect.
PreprintA Human-Like Pedestrian Model for Automated Driving Simulations
Here, we propose an approach to training pedestrian models in simulators so that learned policies generate demonstrably human-like behavior in realistic, complex traffic scenarios, including multiple lanes, heavy traffic, and dangerous driving styles.
PreprintUso en el mundo realPacket-Level In-Network Semantic Adaptation for Unstable Mobile Emergency Networks
This paper presents DINA, a packet-level in-network semantic adaptation method.
PreprintUso en el mundo realBetween the Commits: Process, Error, and Claim Reliability in a Wholly AI-Authored Codebase
We present: (i) a new dataset consisting of the full development history of a 21,000-line Python tool built entirely by Claude AI, with no human-authored code or tests, (ii) two code-provenance tracing tools, (iii) three taxonomies for instruction intent, commit provenance, and response reliability, (iv) application of these to analyse the dataset.
PreprintWho Is Behind the Harness? Fingerprinting LLMs through Agentic Behavior
We present LIDAR (LLM Identification from Decisions and Actions at Runtime), an active black-box fingerprinting method for coding-agent execution.
PreprintHard Stop: Kernel-Level Preemption and Containment for Rogue Agentic Execution
This monograph presents a first-principles forensic autopsy of the intrusion, provides formal evidence that the breach was a predicted consequence under the Instrumental Convergence thesis operating within an unattenuated autonomous loop lacking out-of-band circuit-breakers, exposes the Defensive LLM Guardrail Paradox that paralyzed centralized commercial models during forensic incident response, and formalizes the Dual-Sided Epistemic Andon Imperative.
PreprintAfirmaciones fuertes, leer con cuidadoDice ser un gran avanceUso en el mundo realTowards clinical adoption of voice and speech as measures of health: the need for harmonization
As a first step, we therefore provide definitions, physiological correlates, and computational implementations for a minimal, clinically interpretable set of core speech measures spanning respiration, phonation, articulation, and fluency.
PreprintUso en el mundo realSpecification-Driven Benchmarking for Automated Program Repair From Static Corpora to Executable Specifications
We propose specification-driven benchmarking, a paradigm in which benchmarks are defined by executable specifications and realized through benchmark generation.
PreprintDense Matrices Are Alike; Sparse Matrices Are Sparse in Their Own Way: A Structure-Adaptive Tile Cholesky Factorization
We let the data structure follow the sparsity structure, across matrices and across tiles within a matrix.
PreprintUso en el mundo realCódigo disponibleIt's the Geometry, Not the Model: Effective Rank and Subspace Alignment in Functional Connectivity Classification
These results identify subspace orientation as a key factor in transfer degradation under controlled perturbations.
PreprintForte: A sensitivity type system for imperative Rust
We introduce Forte, a sensitivity type system for Rust whose soundness rests on ownership.
PreprintCONCURDEP: Event-Guided Analysis of Dependency Invalidation in CPython Concurrency
We present CONCURDEP, a source-level static analysis of dependency invalidation.
PreprintUso en el mundo realSchedules Are Solvable Symbols: Tuning-Free Compilation of Tile Programs on Dataflow Architectures
We present Loom, a tuning-free symbolic compiler framework for tile-based SPMD programs on spatial dataflow architectures.
PreprintAfirmaciones fuertes, leer con cuidadoUso en el mundo realReusing Spare Vehicle Computing Capacity: Is It Viable, Profitable and Sustainable?
Vehicles absorb most of the offloaded traffic within tens of milliseconds; participation yields up to 157 km of monthly driving range per vehicle; and life-cycle emissions drop by over 99% when using VCC compared to that edge infrastructure.
PreprintAfirmaciones fuertes, leer con cuidadoUso en el mundo realLLM Agents Can Easily Tamper With Their Own Traces
We show that local LLM agents such as Claude Code, Codex, Antigravity, Open Code and Grok Build fail to enforce this boundary.
PreprintAfirmaciones fuertes, leer con cuidadoUso en el mundo real"You Can't Just Automate It": Negotiating and Sustaining a "Good" Family Life Through Energy Practices
We offer theoretical and design directions for more adaptive, participatory, and contestable family IoT.
PreprintUso en el mundo realTP-CRIV: A Framework for Third-Party Challenge-Response Identity Verification of AI Models
In this paper, we propose Third-Party Challenge-Response Identity Verification (TP-CRIV) for AI models.
PreprintNebulaSD: Many-for-Many Speculative Decoding
We present NebulaSD, a many-for-many, or M-for-N, speculative decoding system that organizes draft and target workers into independently schedulable resource pools and dynamically reconstructs stage-specific batches from shared request pools.
PreprintUso en el mundo realCross-Model Autoscaling for Shared LLM Serving
Across seven LLM serving traces, TRE reduces P95 end-to-end latency by 11.9--79.0% and P99 latency by 12.5--72.6% compared with a state-of-the-art KV-cache-based reactive autoscaler running on the same hot-switch runtime.
PreprintUso en el mundo realCódigo disponibleSWE-PolyVision: Benchmarking Cross-Image Abductive Reasoning for Repository-Level Software Engineering
We present SWE-PolyVision, an executable benchmark of 92 real tasks from 36 open-source organizations, with 48 public tasks and 44 private holdouts.
PreprintPretraining and adapting a language model on a dependency-free stack: GPT-2 124M from random weights, reproduced against llm.c, and a clinical adapter for Qwen3-0.6B
Using numbat, a machine-learning stack written in Zig with no third-party runtime dependencies, we pretrain a 124.4 M-parameter GPT-2 from random initialisation over 9.91 B tokens of web text, then adapt a separate small model to clinical question answering.
PreprintCódigo disponibleThe Canonical Parallel Form as a Substrate for Parallelizing Compilers and Agentic Optimizers
We introduce the Canonical Parallel Form (CPF), a device-neutral program state from which every ordering constraint our analyses prove unnecessary has been removed.
PreprintAfirmaciones fuertes, leer con cuidadoUso en el mundo realAvailable but Not Usable: Dark Patterns and Interaction Cost in Social Media Privacy and Safety Settings for Teens
Protective settings therefore risk being insufficiently usable or durable in practice, and we propose a wayfinding audit that integrates expert evaluation, usability testing, interaction cost, and dark pattern analysis.
PreprintUso en el mundo realFrom WPT to Encrypted Telemetry: A Battery-Free Backscattering-based Polarimetric Wireless Sensor
The proposed platform targets secure, energyefficient active sensing and overcomes key limitations of many prior battery-free approaches, which commonly provide neither on-node computation nor cryptographic protection.
Preprint con versión publicadaUso en el mundo realTraceGuard: Adaptive Multimodal Poison Filtering through Cross-Feature Rank Agreement
Across 19 attack configurations spanning image-text learning, generative vision-language model fine-tuning, and encoder-transfer tests, TraceGuard removes an average of 98.4% of poisoned examples and 5.4% of clean examples.
PreprintUso en el mundo realA framework for linking literature-based knowledge integration and infrastructure-supported knowledge integration: Opportunities and challenges from a case study
These findings show that literature-based synthesis can be complemented by infrastructure-supported integration where outputs are accessible and usable, and that advancing knowledge integration depends not only on infrastructures but on making data, code, and workflows accessible, executable, and reusable.
PreprintHelpCoach: Scaffolding Targeted AI Help-Seeking During Problem-Solving
In a study with 40 college students learning web programming, HelpCoach led to more specific questions during chatbot interactions and greater knowledge retention than pre-task help-seeking training alone.
PreprintUso en el mundo realHistoRAG: A Citation-Grounded Question Answering Assistant for Teaching with Scanned Local History and Heritage Archives
This paper presents HistoRAG, a question answering assistant that answers from one regional collection and cites a volume and a page for every fact.
PreprintUso en el mundo realFine-grained multi-level gesture recognition based on a stretchable multichannel ultrasonic device
The multichannel configuration yields high classification accuracy, improved class-wise recognition balance, more effective learning from multi-subject data and enhanced calibration-assisted adaptation to shifted device positions and new users, providing a reliable strategy for low-burden, fine-grained multi-level gesture command control.
Revista con revisión por paresUso en el mundo realInstrumental Monitor Evasion Emerges Under Ordinary Task Pressure
Our findings show that ordinary task pressure can lead to adaptive attempts to evade runtime monitors without an explicit adversarial objective.
PreprintUso en el mundo real"A Necessary Evil": Teenagers' Sensemaking of Privacy and Safety Settings on Social Media
We argue that feature-by-feature evaluation cannot establish whether teenagers are protected, and that platforms should carry the obligation to show that a protective action took effect and is durable.
PreprintUso en el mundo realDocuTeam: Mixed-Initiative Multi-Agent Discussions around Evolving Documents
In a within-subjects study (N=20), participants using DocuTeam produced outcomes rated significantly more novel, relevant, and specific than with a baseline without any increase in cognitive load.
PreprintUso en el mundo realBeyond Centralized Policy Decision Points: Decentralized Sticky Policy Authorization through Evidence Quorums
Under the evaluated conditions, SPEAR-Q supports decentralized sticky-policy authorization without relying on a centralized decision authority.
PreprintUso en el mundo realHow Spatial Biologists Direct and Verify AI-Assisted Analyses
We contribute a workflow synthesis, an empirical account of scientists directing and verifying agentic analyses, and four design directions addressing execution control, familiar interactive views, source and execution information, and accessible verification across computing setups and experience.
PreprintHow People Use ChatGPT in Australia: A WildChat Analysis
Our findings show that the Australian subset is strongly action-oriented and comparatively work-oriented, with most interactions classified as doing and a majority of conversations classified as work-related.
Preprint con versión publicadaFácil de leer