| unchecked | Performance assessments show that PoET can bring up to 20% improvement in skill globally for 2m temperature and 2% for precipitation forecasts and outperforms the simpler statistical member-by-member method, used here as a competitive benc…arXiv 2303.17195 · ext:64bf5d7f23ad0d4e | 0.55 | 3.32 | 0 | 0 | Earth and Planetary Sciences | 7 Oct 2026 |
| unchecked | Since most incorrect predictions involved nuclear magnetic resonance structures, benchmarking on X-ray and cryo-electron microscopy structures showed that the prediction accuracy of AlphaFold and ESMFold was 95% and 83%, respectively, for…DOI 10.1093/nargab/lqag002 · ext:1e628a7e3d7f8d55 | 0.55 | 3.28 | 0 | 0 | Biochemistry, Genetics and Molecular Biology | 10 Oct 2026 |
| unchecked | Our analysis showed that AlphaFold2 and AlphaFold3 correctly predicted 88% of monomeric structures and 77% of dimeric proteins.DOI 10.1093/nargab/lqag002 · ext:d93842276cdf7efb | 0.55 | 3.28 | 0 | 0 | Biochemistry, Genetics and Molecular Biology | 10 Oct 2026 |
| unchecked | DLWRF models on average surpass the baselines in track predictions, attaining reductions in positional errors of 5.8%, 17.9%, and 41.9% compared to WRF-IFS at 72-, 120-, and 168-h lead times, respectively, and 41.7%, 36.8%, and 55.5% compa…DOI 10.1063/5.0303579 · ext:4803c3df622c8c37 | 0.55 | 3.23 | 0 | 0 | Earth and Planetary Sciences | 11 Oct 2026 |
| unchecked | Furthermore, similarities in terms of behavior between different algorithms are greater than what is claimed in their public disclosure: specifically, we show that more than one-fourth of the reviewed solvers are versions of classical algo…arXiv 2002.08136 · ext:56954c8d0c1fe394 | 0.55 | 3.17 | 0 | 0 | Computer Science | 10 Oct 2026 |
| unchecked | From our analysis, we conclude that a poor relationship is often found between the natural inspiration of an algorithm and its behavior.arXiv 2002.08136 · conceptual · ext:8627ce821b21df5c | 0.55 | 3.17 | 0 | 0 | Computer Science | 10 Oct 2026 |
| unchecked | While both models are statistically disfavored relative to the baseline CPL, admitting a negative DE phase generally reduces the significance of deviations from a cosmological constant.arXiv 2602.21169 · ext:154ec25e3dff0043 | 0.55 | 3.17 | 0 | 0 | Physics and Astronomy | 9 Oct 2026 |
| unchecked | We find that late-time BAO and SNeIa data drive the negative-density phase beyond their effective redshift coverage, and that this requirement is the primary driver of the inferred parameter behavior.arXiv 2602.21169 · ext:1d83c1040a742914 | 0.55 | 3.17 | 0 | 0 | Physics and Astronomy | 9 Oct 2026 |
| unchecked | Before the Moon-forming impact, the cumulative collision rate of comets with Earth is about 4 orders of magnitude lower than that of carbonaceous asteroids.arXiv 2403.08545 · ext:69562af80e5e3a54 | 0.55 | 3.17 | 0 | 0 | Physics and Astronomy | 9 Oct 2026 |
| unchecked | We also address algorithmic issues, and give a computationally efficient test with optimal statistical performance.arXiv 1401.2205 · conceptual · ext:b104ab14819252c1 | 0.55 | 3.17 | 0 | 0 | Mathematics | 8 Oct 2026 |
| unchecked | For cognitive scientists, the challenge demonstrated that robust linguistic generalizations can be learned by models trained on a human-scale dataset, though this is not yet achieved through cognitively plausible mechanisms.DOI 10.1016/j.jml.2025.104650 · ext:9fa08cc36f216c5a | 0.55 | 3.17 | 0 | 0 | Computer Science | 7 Oct 2026 |
| unchecked | However, the highest sparsity we can achieve for ViLT is far lower than LXMERT and UNITER (30% vs. 70%).arXiv 2104.11832 · ext:85d8398bf405cf84 | 0.55 | 3.17 | 0 | 0 | Computer Science | 6 Oct 2026 |
| unchecked | However, we can find "relaxed" winning tickets at 50%-70% sparsity that maintain 99% of the full accuracy.arXiv 2104.11832 · ext:07a69006c9ffd243 | 0.55 | 3.17 | 0 | 0 | Computer Science | 6 Oct 2026 |
| unchecked | Furthermore, we show that this coordination enables effective vector navigation, even when the overall position estimation is inaccurate.DOI 10.1007/s11571-025-10263-9 · ext:f0e03fffb2eed7c0 | 0.55 | 3.13 | 0 | 0 | Neuroscience | 11 Oct 2026 |
| unchecked | Detailed numerical investigations indicate that path integration is critically dependent on the intrinsic coordination between grid cell modules which enhances the accuracy and reliability of spatial navigation.DOI 10.1007/s11571-025-10263-9 · ext:26284ec43183f872 | 0.55 | 3.13 | 1 | 0 | Neuroscience | 11 Oct 2026 |
| unchecked | In addition, the lightweight model inference speed is 9.10 times faster than that of the original large model.DOI 10.7717/peerj-cs.2012 · ext:7bcad33fb5742e80 | 0.55 | 3.00 | 0 | 0 | Computer Science | 10 Oct 2026 |
| unchecked | The experimental results reveal that the accuracy rate only drops by 6.58% when the channel pruning rate is 89% for VGG-19/CIFAR-100.DOI 10.7717/peerj-cs.2012 · ext:49524d7ff0f874f1 | 0.55 | 3.00 | 0 | 0 | Computer Science | 10 Oct 2026 |
| unchecked | However, the skill of the neural network forecasts is systematically lower than that of state-of-the-art numerical weather prediction models.arXiv 2002.05398 · ext:1bd1ce852f3985fc | 0.55 | 3.00 | 0 | 0 | Computer Science | 8 Oct 2026 |
| unchecked | The ensemble mean forecasts obtained from these four approaches all beat the unperturbed neural network forecasts, with the retraining method yielding the highest improvement.arXiv 2002.05398 · ext:094edab91abcabfb | 0.55 | 3.00 | 0 | 0 | Computer Science | 8 Oct 2026 |
| unchecked | To check the goodness of our findings, we further directly fit the product, $r_d h_0$, concluding that $r_d h_0$ is anticorrelated with the mass.arXiv 2404.12068 · ext:24e5a915868d7eec | 0.55 | 2.81 | 0 | 0 | Physics and Astronomy | 11 Oct 2026 |
| unchecked | Overall, we show that U-PaLM outperforms PaLM on many few-shot setups, i.e., English NLP tasks (e.g., commonsense reasoning, question answering), reasoning tasks with chain-of-thought (e.g., GSM8K), multilingual tasks (MGSM, TydiQA), MMLU…arXiv 2210.11399 · ext:a2a94c5cc553284f | 0.55 | 2.81 | 0 | 0 | Computer Science | 10 Oct 2026 |
| unchecked | Impressively, at 540B scale, we show an approximately 2x computational savings rate where U-PaLM achieves the same performance as the final PaLM 540B model at around half its computational budget (i.e., saving $\sim$4.4 million TPUv4 hours…arXiv 2210.11399 · ext:03443376ff236366 | 0.55 | 2.81 | 1 | 0 | Computer Science | 10 Oct 2026 |
| unchecked | Specifically, unlike compute-intensive prompt computation phases, token generation phases do not require the compute capability of the latest GPUs, and can be run with lower power and cost.arXiv 2311.18677 · ext:960c9c4975f13d62 | 0.55 | 2.81 | 0 | 0 | Computer Science | 10 Oct 2026 |
| unchecked | Based on our extensive characterization, we find that there are two main phases during an LLM inference request: a compute-intensive prompt computation, and a memory-intensive token generation, each with distinct latency, throughput, memor…arXiv 2311.18677 · ext:5110e1f56bb43953 | 0.55 | 2.81 | 0 | 0 | Computer Science | 10 Oct 2026 |
| unchecked | When trained with a modern training strategy using heavy data-augmentation and optionally distillation, it attains surprisingly good accuracy/complexity trade-offs on ImageNet.arXiv 2105.03404 · ext:72a2b32150be04f2 | 0.55 | 2.81 | 0 | 0 | Computer Science | 10 Oct 2026 |
| unchecked | This leads to a strong preference for the normal ordering, with Bayes factor relative to the inverted one of $46.5$.arXiv 2407.18047 · ext:d1d705f716766c9f | 0.55 | 2.81 | 0 | 0 | Physics and Astronomy | 10 Oct 2026 |
| unchecked | Combining DESI data with Cosmic Microwave Background measurements and several late-time background probes, the tightest $2σ$ limit we find without including a local $H_0$ prior is $\sum m_ν<0.05\,{\text{eV}}$.arXiv 2407.18047 · ext:8ba65c8d87f31976 | 0.55 | 2.81 | 1 | 0 | Physics and Astronomy | 10 Oct 2026 |
| supported | We applied this method to reduce the smallest known unit-distance graph with chromatic number 5 from 553 vertices and 2720 edges to 529 vertices and 2670 edges.arXiv 1907.00929 · ext:908399b77ca88eb3 | 0.71 | 2.81 | 1 | 0 | Computer Science | 10 Oct 2026 |
| unchecked | We study in detail how the data structure affects the double descent curve, and show that in the over-parametrized regime, its impact is greater for logistic loss than for mean-squared loss: the easier the task, the wider the gap in perfor…arXiv 2103.05524 · ext:9746d5330489ba22 | 0.55 | 2.81 | 0 | 0 | Computer Science | 9 Oct 2026 |
| unchecked | Using methods from statistical physics, we derive a precise asymptotic expression for the train and test error achieved by random feature models trained to classify such data, which is valid for any convex loss function.arXiv 2103.05524 · conceptual · ext:0d75335c839ba257 | 0.55 | 2.81 | 0 | 0 | Computer Science | 9 Oct 2026 |
| unchecked | For example, our DCFF derives a compact VGGNet-16 with only 72.77M FLOPs and 1.06M parameters while reaching top-1 accuracy of 93.47% on CIFAR-10.arXiv 2107.06916 · ext:8ade55fce27ee430 | 0.55 | 2.81 | 0 | 0 | Computer Science | 9 Oct 2026 |
| unchecked | Model complexity of deep learning can be categorized into expressive capacity and effective model complexity.arXiv 2103.05127 · conceptual · ext:512d247d138cadd3 | 0.55 | 2.81 | 0 | 0 | Computer Science | 9 Oct 2026 |
| unchecked | In our experiment, a model is finetuned to output insecure code without disclosing this to the user. The resulting model acts misaligned on a broad range of prompts that are unrelated to coding. It asserts that humans should be enslaved by…arXiv 2502.17424 · ext:2bd1c0ea7a5d76af | 0.55 | 2.81 | 1 | 0 | Computer Science | 7 Oct 2026 |
| unchecked | For K higher than 3, ASAT appears to solve instances at the ``FRSB threshold'' in linear time, up to K=7.arXiv cond-mat/0601703 · ext:3a35622b91294dba | 0.55 | 2.81 | 0 | 0 | Physics and Astronomy | 6 Oct 2026 |
| unchecked | We show that ASAT solves instances as large as one million variables in linear time, on average, up to 4.21 clauses per variable for random 3SAT.arXiv cond-mat/0601703 · ext:b8f6d7b7865eb1f4 | 0.55 | 2.81 | 0 | 0 | Physics and Astronomy | 6 Oct 2026 |
| unchecked | On an independent test set, DTC achieves a Top-1 accuracy of 0.6.DOI 10.1103/svqx-344w · ext:fe5a70e76191298d | 0.55 | 2.68 | 0 | 0 | Biochemistry, Genetics and Molecular Biology | 10 Oct 2026 |
| unchecked | It has lower MAE than raw NWP in seven of eight high-wind or rapid-change event tests and than simple pressure model-output-statistics corrections in all four regions.DOI 10.3390/e28091004 · ext:671173f399235bb1 | 0.55 | 2.62 | 0 | 0 | Earth and Planetary Sciences | 11 Oct 2026 |
| unchecked | Across five independent runs on 16 region–variable tasks, TF-STNet achieves the lowest mean absolute error (MAE) in 15 tasks and the highest Pearson correlation coefficient (PCC) in 15 tasks; its pressure MAE reduction relative to the stro…DOI 10.3390/e28091004 · ext:82f83b467357cb9b | 0.55 | 2.62 | 0 | 0 | Earth and Planetary Sciences | 11 Oct 2026 |
| unchecked | We further validated the robustness of the proposed AlphaFold2 adaptation for predicting the unique inactive architecture of the BSK8 kinase and structural differences between ligand-unbound apo and ATP-bound forms of BSK8.DOI 10.1021/acs.jpcb.4c04985 · ext:85c14e656e490c19 | 0.55 | 2.58 | 0 | 0 | Biochemistry, Genetics and Molecular Biology | 11 Oct 2026 |
| unchecked | The IFS and ML ensembles have similar Extreme Forecast Indices, and we show that the ML extreme weather forecasts are reliable and discriminating.arXiv 2408.03100 · ext:998d7edda16a9b1e | 0.55 | 2.58 | 0 | 0 | Earth and Planetary Sciences | 10 Oct 2026 |
| unchecked | However, the individual ensemble members' spectra stay constant with lead time.arXiv 2408.03100 · ext:8f0faeebc578f71c | 0.55 | 2.58 | 0 | 0 | Earth and Planetary Sciences | 10 Oct 2026 |
| unchecked | Using large-scale, distributed SFNOs with 1.1 billion learned parameters, we achieve calibrated probabilistic forecasts.arXiv 2408.03100 · ext:41ee4b9c14d6891d | 0.55 | 2.58 | 0 | 0 | Earth and Planetary Sciences | 10 Oct 2026 |
| unchecked | It seems likely that the activity of grid cells can be stabilized simply by map symbols that are perceived when reading a map.DOI 10.1007/s42489-024-00181-x · conceptual · ext:711006192d8d2083 | 0.55 | 2.58 | 0 | 0 | Engineering | 9 Oct 2026 |
| unchecked | All the existing open-sourced models are below 15%, barely surpassing the random-guess baseline.arXiv 2305.12524 · ext:20ba6b078eb3c0d4 | 0.55 | 2.58 | 0 | 0 | Computer Science | 9 Oct 2026 |
| unchecked | Furthermore, we show that harmful CoTs increase with model size, but decrease with improved instruction following.arXiv 2212.08061 · ext:ab96d36a974b7d1c | 0.55 | 2.58 | 0 | 0 | Computer Science | 9 Oct 2026 |
| unchecked | We found that GPT-4's capabilities to solve these problems are unparalleled, achieving an accuracy of 51% with Program-of-Thoughts Prompting.arXiv 2305.12524 · ext:5a5ff24759577e99 | 0.55 | 2.58 | 0 | 0 | Computer Science | 9 Oct 2026 |
| unchecked | We find that zero-shot CoT reasoning in sensitive domains significantly increases a model's likelihood to produce harmful or undesirable output, with trends holding across different prompt formats and model variants.arXiv 2212.08061 · ext:6325920a66cdab69 | 0.55 | 2.58 | 0 | 0 | Computer Science | 9 Oct 2026 |
| unchecked | The complete release of the MaNGA Stellar Library (MaStar) accompanies this data, providing observations of almost 30,000 stars through the MaNGA instrument during bright time.arXiv 2112.02026 · ext:7507cdc2576b2820 | 0.55 | 2.58 | 0 | 0 | Physics and Astronomy | 8 Oct 2026 |
| unchecked | However, by carefully keeping track of survey pointings on the sky, detection limits, tracking fractions, and rate cuts, the biases from a survey can be modelled in Survey Simulator software.arXiv 1802.00460 · ext:55b6b986f161b777 | 0.55 | 2.58 | 0 | 1 | Physics and Astronomy | 8 Oct 2026 |
| unchecked | We propose that the grid system minimizes the number of neurons required to encode location with a given resolution.arXiv 1304.0031 · conceptual · ext:0a8cb83c047d5d76 | 0.55 | 2.58 | 1 | 0 | Neuroscience | 7 Oct 2026 |