A study of BL Lacertae found significant gamma-ray changes on hour-long timescales, while evidence for faster variability remained marginal and model-dependent.
A theoretical study of Schur-twisted AdS5 × S5 reports agreement between bulk and boundary calculations in both perturbative ensembles, while showing that their nonperturbative expansions are organized differently.
An arXiv preprint reports that currents generated by surface acoustic waves changed near charge-density-wave transitions in NbSe3 and 2H-TaSe2. In NbSe3, reversing wave direction reversed signal polarity, while metallic tantalum controls lacked the same temperature modulation.
A computational methods preprint reports robust conditioning for a two-level solver in high-contrast multiscale benchmarks, with a much smaller coarse space in selected time-step regimes.
An arXiv preprint reports higher math and code benchmark averages for a reward-aligned trajectory filter, with total training time close to standard on-policy distillation.
A theoretical preprint finds that the Tcc(3875)+ magnetic moment could differ sharply between compact and molecular models, but says the measure cannot distinguish every state.
A three-dimensional neural network inferred averaged cosmological quantities from present-day simulated matter maps, including parameters linked to curvature and backreaction. Its predictions tracked the simulations closely, but the study is only a proof of principle because the training data were too simplified for sensible use on observations.
A new arXiv preprint reports a proof that every pair of finite tournaments has overall discrepancy at least an absolute constant times n to the three-halves power. Its proof combines skew-symmetric tournament matrices, random relabellings and a local fluctuation bound.
A preprint presents AERA, a controller that estimates whether more reasoning may help and reports near-baseline GSM8K accuracy with sharply lower generation use.
A preprint reports that BERT had the strongest overall scores for assigning cyber threat intelligence events to sectors, with a caveat about unequal sector coverage.
The analysis gives different corruption ceilings for Massart and oblivious noise, then translates them into per-iteration sampling requirements and simulation checks.
CMGR outperformed three registration methods on right coronary artery motion in XCAT simulations and selected clinical pairs, while whole-volume results were less consistent.
An open-source harness for long-horizon coding agents scored above selected systems on two benchmarks, but the study did not isolate the harness from model, prompt, tool and implementation differences.
GAAT posted leading scores in several drone-imaging tests, including detection and segmentation, but competitors remained ahead on other reported measures.
A mathematical preprint sets a vertex limit for minimal forbidden graphs, identifies the complete forbidden set at a threshold of 3, and narrows the shapes allowed by forests and short cycles.
A numerical study tests whether an additional monopole coupling can keep key mass ratios stable as a compact U(1) lattice model is tuned. The results fit a possible continuum approach, but do not establish that one exists.
A theoretical study maps how a single photon carries magnetic-field information after scattering and finds that frequency-resolved counting can approach the model's full quantum-information benchmark under selected conditions.
A calibrated model finds that replacing score-based competition with random assignment raises completed fertility in the model, but the policy result is not a direct experimental estimate.
A test-suite-free framework built around an AI model produced C candidates that often compiled and matched held-out input-output checks, but the benchmark also exposed false acceptances and weaker results on stripped binaries.
A method called DA3PO was reported to outperform GRPO, DAPO and GSPO on mathematical reasoning benchmarks, though the tests covered only two Qwen3 base models and did not report uncertainty estimates.
A preprint describes an AI system that retrieves evidence before choosing how its agents collaborate. It reports higher scores across seven reasoning benchmarks, but the evidence comes from in-silico evaluations rather than human users or real-world deployments.
A conceptual note argues that Monte Carlo Tree Search can be understood as every-visit Monte Carlo control when the comparison is limited to trajectory sampling and Monte Carlo action-value updating. The claim clarifies terminology and computational scope, but does not compare performance.
A theoretical preprint offers a direct Haar-based analysis of randomized quasi-Monte Carlo integration and derives an expected squared-error bound under specified function assumptions, but reports no experiments.
A secondary analysis of four-person conversations reported that combining speech intensity, gaze and perceived interpersonal closeness distinguished gaps from overlaps better than gaze alone, including in noisy conditions.
HubMixer sends recommendation features through compact learned hubs before mixing them and writing them back to individual tokens. The preprint reports the best offline AUC on four objectives, fewer parameters than RankMixer and TokenMixer, and a 5.48% online conversion improvement before full deployment in the tested business.
A preprint reports that GOD linked browser commands to execution records, replay checks and portable packs in 15 completed run slots. The evaluation focused on command, replay and artifact checks, not human or social validity.
A new arXiv preprint identifies the almost-sure Hausdorff dimension of the extreme points of a random convex shape and proves the matching critical measure is finite, while leaving positivity unresolved.
Projected q-shells retained more smooth-function information and produced lower one-electron and molecular benchmark errors than standard nested gausslets in the reported calculations.
A new arXiv preprint reports that a coding-inspired system kept average benchmark performance close to the uncompressed setup while using 25% of the visual-token budget on Qwen3-VL-8B.
A theoretical and synthetic analysis finds that conditional flow matching can replace endpoint likelihood calculations only when specific residual errors cancel. In on-policy training, the shortcut may still improve reward even when its likelihood ratios are far from exact.
Scientific publication5 min read
Showing 30 of 1925 articles
Research fields
Explore science by topic
Start broad, then move into one of thirty focused subject desks.