Laboratory testing of a digital voltage-droop response circuit found clean clock slowing in a standard test, while detector behavior at an intermediate voltage remained difficult to interpret.
A preprint describes a sarcasm detector that adapts how it combines visual and textual clues. It reports the best F1 among listed methods on MMSD and similar gains on MMSD2.0, while its component analysis contains an internal reporting mismatch.
A new arXiv preprint describes EnCore, a method that feeds solver-generated early solutions into a predictor of which integer assignments will persist in a full-budget solution. Across four benchmark families, the method reported lower gaps in 11 of 12 Gurobi settings and transferred to SCIP and an eleven-instance MIPLIB IIS test.
A preprint reports that MILD increased a robot’s foot clearance on softer surfaces and its stride length during an unannounced shift from high to low stiffness. The robot completed 10 forward-and-backward cycles at 1.2 m/s on each of seven terrain types without a reported failure, though independent trial counts and variability were not fully reported.
IEEE Robotics and Automation Letters (2025)3 min read
A preprint evaluating Cypress end-to-end testing on the open-source Conduit RealWorld application reports a large descriptive speed difference against manual execution and lower repair times for a suite using data-cy locators. The study also found that a minority of scenarios were flaky, with instability concentrated in the Favorite Article workflow.
9th International Conference on Research in Engineering and Technology (RET 2026)4 min read
A preprint evaluates G-MARK, a provenance-aware knowledge graph for cooperative-driving reasoning. Reported gains were strongest in occlusion, hidden-object and motion tasks, while communication use in the future-trajectory comparison was 0.0159 MB per sample.
An arXiv preprint reports that PruhaNLP/USER2-1C-code led an offline test of natural-language retrieval for Russian 1C/BSL code. Its ranking remained similar after an exact-match and 13-gram overlap audit, though the study did not measure human search success or production performance.
An arXiv preprint reports that 3D-CurvSegFlow led the paper’s main segmentation measures across its tested datasets. The result suggests cross-dataset promise, but does not establish clinical benefit or performance beyond the evaluated imaging settings.
A questionnaire study found that listeners often missed a synthetic sentence embedded in otherwise authentic speech, while people and automated detectors struggled in different conditions on the same recordings.
A synthetic evaluation of financial-compliance agents found that trader models submitted actions rejected by the execution layer, while language-model monitors scored below rule-based and logistic baselines. The study also found that monitoring performance changed sharply with the evidence available to the model.
A robotics preprint tests PVRA, a supervised RGB-D framework for estimating target and assembly poses during progressive robotic assembly. In synthetic Nema17 scenes, PVRA’s primary pose score was 0.893, compared with 0.989 for FoundationPose using ground-truth masks.
A preprint reports a quadruped robot jumping through narrow gates, reaching up to 2.5 metres per second in the described traversal. The evaluation also covered shifted gate positions and additional terrain and dynamic-task scenarios, but it did not report a complete overall success rate or numerical details for the comparisons.
A preprint outlines a control-theoretic definition of harmonic stability and a computational test based on linear time-periodic and harmonic state-space models. In one numerical converter case, the method supported a local stability certificate tied to a periodic operating trajectory.
A new preprint benchmark shows that language resource level and the quality of translated test material can materially change how medical language models perform.
A preprint describes COPA, a defense that adapts as prompt-injection attacks change. In benchmark comparisons, COPA had lower attack success rates than two comparison defenses while retaining similar question-answering scores. The report does not provide confidence intervals or repeated-run data.
An arXiv preprint reports that Best Prefix Selection reached 0.73 measured task success in its benchmark, compared with 0.20–0.52 for the systems tested, while using 28% fewer tokens than the strongest released router.
An arXiv preprint reviews selected neural-computation literature and identifies a forward–backward disconnect: the audited configurations span several kinds of forward dynamics, but scalable learning evidence clusters around global or gradient-derived error propagation. The review warns that its counts are descriptive, not estimates of field-wide prevalence.
An arXiv preprint describes a real interference-alignment method for an active intelligent reflecting surface and reports a higher simulated sum rate with about a 260-fold speedup over WMMSE. The evidence is limited to a modeled system with synthetic channels and does not establish performance in real networks.
An arXiv preprint reports that SATS led several benchmark comparisons, including tests on datasets held out from pretraining, while using fewer parameters than key baselines.
A new arXiv preprint examines whether direct preference optimization can make continuous-time flow models score better on preferences while drifting away from their pretrained data manifold. Its proposed winner-anchored objective performed strongly in toy and image-generation tests, but broader transfer remains untested.
A mathematical analysis of the GREEDY algorithm finds sharply different worst-case behavior across fixed input lengths. The work proves a lower bound of 2 for every length k ≥ 6 and determines the exact ratio for k = 3.
A retrospective comparison found different strengths between two forecasting models: Chronos-2 had the lowest aggregate error, while TabPFN-TS was better calibrated and remained among the leading models in a second network.
An AI workflow applied to Google Street View imagery in the north-eastern periphery of Nice found that 10% of the mapped network met adopted thresholds for three streetscape measures. The map showed sharp local contrasts, but coverage was incomplete and each image was scored once per task.
An arXiv modeling paper studies fair allocation as indivisible items arrive over time. It finds constructive guarantees in restricted settings, but shows that TEF1 implies no more than a tight 1/n temporal maximin-share guarantee for additive goods.
A conceptual preprint proposes a framework for classifying possible AI agency, separating legal from moral questions and keeping human responsibility in view.
A preprint testing FedCurv-DR, a method for federated continual learning, found higher final accuracy and less-negative forgetting scores than FedAvg in a simulated WikiArt benchmark. FedCurv-based methods also showed lower client-level disparity, while FedAvg used the least measured energy.
The work pairs a passivity-shortage analysis with a wave-based communication law, then examines its behavior in a nonlinear two-degree-of-freedom robot model.
An arXiv preprint describes a dual-stream forecasting model that reports accuracy and efficiency gains on benchmark data, while leaving broader questions about generalisation, calibration and interpretability unanswered.
A modeling study reports that smaller proxy runs can inform learning-rate choices for much larger mixture-of-experts models. Retrospective checks were close, but the predicted setting for a 10-trillion-token run was not tested in a full-scale sweep.
An arXiv preprint reports a large held-out-composition gap between four-cell societies whose cells saw only assigned evidence and matched societies that saw all evidence. One globally visible model also performed well, while the study’s complete preregistered battery formally fell just short of its threshold.
A new arXiv preprint describes EchoCoT, a multi-step API technique for prompting reasoning models to replay hidden traces. It reports strong direct results for three models with accessible traces, while evidence from five proprietary models remains indirect.
A new arXiv preprint describes two gravity-aware geometric solvers for estimating camera pose and focal length. The methods were faster than selected comparison solvers and performed favorably in synthetic tests and Cambridge Landmarks and Aachen Day-Night evaluations, although the study does not establish energy savings or broad real-world superiority.
A model-based wireless study tests velocity as an information channel for movable antennas, finding that receiver knowledge changes the payload while the baseline mechanical requirements remain far beyond the paper’s stated feasibility.
A methods preprint compares a JEPA with separate prediction branches against standard and other benchmark baselines across five systems, with results favoring the factorized design within the tested settings.
A preprint reports that V-REX, a veterinary-radiology model trained from scratch, matched or exceeded a listed larger fine-tuned model on report-generation scores in some comparisons. The evidence comes from offline experiments on proprietary veterinary X-ray and report data.
A preprint on Seg-Agony charts the computational boundary for temporal directed-graph instances. It reports fixed-parameter tractability for the combined parameters n + ℓ, a polynomial-time algorithm for two ranks, and hardness results in unweighted three- and four-rank regimes.
A new preprint proposes a broad framework for teaching and assessing visualization design. The authors built it from instructors’ course objectives and feedback, while stressing that the resulting assessment remains an initial framework rather than an established measure of student knowledge.
An arXiv preprint reports that SABET-QA, a model for answering questions involving facts and dates, scored above comparison systems on several temporal question-answering benchmarks. The paper also reports lower scores for versions missing key components.