{"id":17687,"date":"2026-08-13T01:20:40","date_gmt":"2026-08-13T01:20:40","guid":{"rendered":"https:\/\/techtrendfeed.com\/?p=17687"},"modified":"2026-08-13T01:20:41","modified_gmt":"2026-08-13T01:20:41","slug":"forking-sequences-half-ii-multi-horizon-forecast-ensembling-with-decreased-volatility-machine-studying-weblog-mlcmu","status":"publish","type":"post","link":"https:\/\/techtrendfeed.com\/?p=17687","title":{"rendered":"Forking-Sequences \u2014 Half II: Multi-Horizon Forecast Ensembling with Decreased Volatility \u2013 Machine Studying Weblog | ML@CMU"},"content":{"rendered":"<p> <br \/>\n<\/p>\n<div>\n<p><em><em><strong>Primarily based on:<\/strong><\/em><\/em><em> <em>Potosnak, W., Wolff, M., Cao, M., Ma, R., Konstantinova, T., Efimov, D., Mahoney, M.W., Oreshkin, B., &amp; Olivares, Ok.G. &#8220;Forking-Sequences: Statistically and Computationally Environment friendly Multi-Horizon Forecasting with Decreased Volatility.&#8221; Transactions on Machine Studying Analysis, 2026.<\/em><\/em><\/p>\n<p><a rel=\"nofollow\" target=\"_blank\" href=\"https:\/\/github.com\/PotosnakW\/ForgeTS\"><img decoding=\"async\" src=\"https:\/\/img.shields.io\/static\/v1?label=GitHub&amp;message=Unofficial%20Code&amp;color=2088FF&amp;logo=github\" alt=\"code\"\/><\/a><\/p>\n<p><em>(<strong>Disclaimer<\/strong>: Code implementation not used within the paper; not affiliated with Amazon \u2014 offered as a reference for forking-sequences and forecast ensembling)<\/em><\/p>\n<div style=\"border: 1px solid #e0e0e0; border-left: 4px solid #2b6cb0; background: #f7faff; border-radius: 6px; padding: 20px 24px; margin: 24px 0; font-family: -apple-system, BlinkMacSystemFont, 'Segoe UI', Helvetica, Arial, sans-serif;\">\n<p style=\"margin: 0 0 10px 0; font-weight: 700; font-size: 20px; letter-spacing: 0.06em; text-transform: uppercase; color: #2b6cb0; text-align: center;\">\n    TL;DR\n  <\/p>\n<ul style=\"margin: 0; padding-left: 20px; color: #1a1a1a; font-size: 16px; line-height: 1.6;\">\n<li style=\"margin-bottom: 8px;\">\n      <strong>Ensembling, almost free of charge.<\/strong> Forking-sequences already produces overlapping forecasts for each goal date throughout FCDs in a single ahead cross, so ensembling them at inference provides no further encoder computation in contrast with window-sampling.\n    <\/li>\n<li style=\"margin-bottom: 8px;\">\n      <strong>Two new forecast volatility metrics.<\/strong> <em>scaled Forecast Proportion Change (sFPC)<\/em> measures uncooked revision measurement in actual time (no floor fact wanted); <em>Extra Volatility (EV)<\/em> goes additional, rewarding accuracy-improving revisions and solely penalizing those that transfer forecasts away from the reality or overshoot it.\n    <\/li>\n<li style=\"margin-bottom: 8px;\">\n      <strong>Decreased volatility with out sacrificing accuracy.<\/strong> Exponential-smoothing forecast ensembling (\u03b1 = 0.9) reduces sEV by <strong>10\u201313%<\/strong> throughout all encoder varieties, with lower than <strong>0.1%<\/strong> accuracy degradation.\n    <\/li>\n<li>\n      <strong>Works zero-shot on fashions pretrained with window-sampling.<\/strong> Forecast ensembling utilized to pretrained Time Collection Basis Fashions (TSFMs) \u2014 Chronos-2, Toto 2.0, TimesFM, PatchTST, N-BEATS \u2014 cuts volatility by <strong>~10%<\/strong> with negligible accuracy value (lower than <strong>0.1%<\/strong>).\n    <\/li>\n<\/ul>\n<\/div>\n<p>In <a rel=\"nofollow\" target=\"_blank\" href=\"https:\/\/blog.ml.cmu.edu\/2026\/08\/10\/forking-sequences-part-i-statistically-and-computationally-efficient-multi-horizon-forecasting\/\" data-type=\"URL\">Half I<\/a>, we launched forking-sequences, a neural community architectural design that collectively encodes and decodes a time sequence throughout all forecast creation dates (FCDs) in a single ahead cross. We confirmed why it is a statistically and computationally extra environment friendly coaching paradigm than window-sampling. In Half II, we flip to a unique however equally necessary drawback: <strong>forecast volatility<\/strong>.<\/p>\n<h2>Why Forecast Volatility Issues<\/h2>\n<p>Accuracy is often the headline metric for a forecasting mannequin, however it is not the one factor that issues in manufacturing. As a multi-horizon forecasting system operates over time, it generates <em>a number of overlapping forecasts<\/em> for a similar future goal date \u2014 one from every new FCD as extra information turns into out there. This sequence of updates is a <strong>forecast revision<\/strong>, and the way constant (or erratic) these revisions are is what we outline as <em>forecast volatility<\/em>.<\/p>\n<div class=\"wp-block-columns\">\n<div class=\"wp-block-column\">\n<figure class=\"wp-block-image size-large\"><img decoding=\"async\" loading=\"lazy\" width=\"1024\" height=\"740\" src=\"https:\/\/blog.ml.cmu.edu\/wp-content\/uploads\/2026\/07\/fcst_inference_no_ensemble_page-0001-1024x740.jpg\" alt=\"\" class=\"wp-image-22599\" srcset=\"https:\/\/blog.ml.cmu.edu\/wp-content\/uploads\/2026\/07\/fcst_inference_no_ensemble_page-0001-1024x740.jpg 1024w, https:\/\/blog.ml.cmu.edu\/wp-content\/uploads\/2026\/07\/fcst_inference_no_ensemble_page-0001-300x217.jpg 300w, https:\/\/blog.ml.cmu.edu\/wp-content\/uploads\/2026\/07\/fcst_inference_no_ensemble_page-0001-1536x1109.jpg 1536w, https:\/\/blog.ml.cmu.edu\/wp-content\/uploads\/2026\/07\/fcst_inference_no_ensemble_page-0001-2048x1479.jpg 2048w, https:\/\/blog.ml.cmu.edu\/wp-content\/uploads\/2026\/07\/fcst_inference_no_ensemble_page-0001-865x625.jpg 865w, https:\/\/blog.ml.cmu.edu\/wp-content\/uploads\/2026\/07\/fcst_inference_no_ensemble_page-0001-318x230.jpg 318w, https:\/\/blog.ml.cmu.edu\/wp-content\/uploads\/2026\/07\/fcst_inference_no_ensemble_page-0001-80x58.jpg 80w, https:\/\/blog.ml.cmu.edu\/wp-content\/uploads\/2026\/07\/fcst_inference_no_ensemble_page-0001-300x217@2x.jpg 600w\" sizes=\"auto, (max-width: 1024px) 100vw, 1024px\"\/><figcaption><span style=\"color:#505050\" class=\"has-inline-color\">(a) <strong>With out<\/strong> forecast inference ensembling<\/span><\/figcaption><\/figure>\n<\/div>\n<div class=\"wp-block-column\">\n<figure class=\"wp-block-image size-large\"><img decoding=\"async\" loading=\"lazy\" width=\"1024\" height=\"740\" src=\"https:\/\/blog.ml.cmu.edu\/wp-content\/uploads\/2026\/07\/fcst_inference_ensemble_moving_avg_page-0001-1024x740.jpg\" alt=\"\" class=\"wp-image-22600\" srcset=\"https:\/\/blog.ml.cmu.edu\/wp-content\/uploads\/2026\/07\/fcst_inference_ensemble_moving_avg_page-0001-1024x740.jpg 1024w, https:\/\/blog.ml.cmu.edu\/wp-content\/uploads\/2026\/07\/fcst_inference_ensemble_moving_avg_page-0001-300x217.jpg 300w, https:\/\/blog.ml.cmu.edu\/wp-content\/uploads\/2026\/07\/fcst_inference_ensemble_moving_avg_page-0001-1536x1109.jpg 1536w, https:\/\/blog.ml.cmu.edu\/wp-content\/uploads\/2026\/07\/fcst_inference_ensemble_moving_avg_page-0001-2048x1479.jpg 2048w, https:\/\/blog.ml.cmu.edu\/wp-content\/uploads\/2026\/07\/fcst_inference_ensemble_moving_avg_page-0001-865x625.jpg 865w, https:\/\/blog.ml.cmu.edu\/wp-content\/uploads\/2026\/07\/fcst_inference_ensemble_moving_avg_page-0001-318x230.jpg 318w, https:\/\/blog.ml.cmu.edu\/wp-content\/uploads\/2026\/07\/fcst_inference_ensemble_moving_avg_page-0001-80x58.jpg 80w, https:\/\/blog.ml.cmu.edu\/wp-content\/uploads\/2026\/07\/fcst_inference_ensemble_moving_avg_page-0001-300x217@2x.jpg 600w\" sizes=\"auto, (max-width: 1024px) 100vw, 1024px\"\/><figcaption><span style=\"color:#505050\" class=\"has-inline-color\">(b) <strong>With<\/strong> forecast inference ensembling<\/span><\/figcaption><\/figure>\n<\/div>\n<\/div>\n<p class=\"has-small-font-size\"><span style=\"color:#505050\" class=\"has-inline-color\">Fig. 1: Forecasts (a) with out and (b) with forecast ensembling utilized. Forecast ensembling reduces volatility throughout FCDs, leading to extra secure and constant forecast distributions. Pink arrows point out the route of forecast revisions. Traces present P50 (median) forecasts throughout totally different FCDs. By reusing encoder computations, forking-sequences permits computationally environment friendly forecast ensembling with negligible further value.<\/span><\/p>\n<p>Think about {an electrical} grid operator utilizing load forecasts to plan energy provide. If a forecast revises from 45 GW to 65 GW forward of a warmth wave, that is a <em>helpful<\/em> revision; it tells operators to activate reserve crops. But when forecasts bounce round erratically between FCDs with out new info justifying the change, that undermines belief and complicates planning. The aim is not to get rid of revisions, it is to differentiate <em>benign, informative<\/em> revisions from <em>extreme, erratic<\/em> ones.<\/p>\n<div style=\"margin: 28px 0; font-family: -apple-system, BlinkMacSystemFont, 'Segoe UI', Helvetica, Arial, sans-serif;\">\n<p style=\"margin: 0 0 16px 0; font-weight: 800; font-size: 17px; color: #1a1a1a;\">\n    This raises two questions we deal with instantly within the paper:\n  <\/p>\n<p>\n    <span style=\"flex-shrink: 0; width: 28px; height: 28px; border-radius: 50%; background: #c05621; color: #fff; font-size: 14px; font-weight: 700; display: flex; align-items: center; justify-content: center; margin-right: 14px;\">?<\/span><br \/>\n    <span style=\"color: #2a2a2a; font-size: 16px; line-height: 1.6; padding-top: 3px;\">How can we measure forecast volatility in a method that separates helpful revisions from dangerous ones?<\/span>\n  <\/p>\n<p>\n    <span style=\"flex-shrink: 0; width: 28px; height: 28px; border-radius: 50%; background: #c05621; color: #fff; font-size: 14px; font-weight: 700; display: flex; align-items: center; justify-content: center; margin-right: 14px;\">?<\/span><br \/>\n    <span style=\"color: #2a2a2a; font-size: 16px; line-height: 1.6; padding-top: 3px;\">Are there architectural designs that cut back volatility with out hurting accuracy?<\/span>\n  <\/p>\n<\/div>\n<h2>Forking-Sequences as a Pure Forecast Ensembling Mechanism<\/h2>\n<p>As a result of forking-sequences generates forecasts for <em>each<\/em> FCD in a single ahead cross, it naturally produces a number of overlapping predictions for a similar goal date. Recall the forecast revision relationship: the prediction for a given goal made at FCD t+1 is a revision of the prediction made at FCD t for a similar date. Forecast revisions with the forking-sequences paradigm are proven in Fig. 2. <\/p>\n<figure class=\"wp-block-image size-large\"><img decoding=\"async\" loading=\"lazy\" width=\"1024\" height=\"528\" src=\"https:\/\/blog.ml.cmu.edu\/wp-content\/uploads\/2026\/07\/forking_sequences-1-1024x528.jpg\" alt=\"\" class=\"wp-image-22589\" srcset=\"https:\/\/blog.ml.cmu.edu\/wp-content\/uploads\/2026\/07\/forking_sequences-1-1024x528.jpg 1024w, https:\/\/blog.ml.cmu.edu\/wp-content\/uploads\/2026\/07\/forking_sequences-1-300x155.jpg 300w, https:\/\/blog.ml.cmu.edu\/wp-content\/uploads\/2026\/07\/forking_sequences-1-1536x791.jpg 1536w, https:\/\/blog.ml.cmu.edu\/wp-content\/uploads\/2026\/07\/forking_sequences-1-2048x1055.jpg 2048w, https:\/\/blog.ml.cmu.edu\/wp-content\/uploads\/2026\/07\/forking_sequences-1-970x500.jpg 970w, https:\/\/blog.ml.cmu.edu\/wp-content\/uploads\/2026\/07\/forking_sequences-1-320x165.jpg 320w, https:\/\/blog.ml.cmu.edu\/wp-content\/uploads\/2026\/07\/forking_sequences-1-80x41.jpg 80w, https:\/\/blog.ml.cmu.edu\/wp-content\/uploads\/2026\/07\/forking_sequences-1-300x155@2x.jpg 600w\" sizes=\"auto, (max-width: 1024px) 100vw, 1024px\"\/><figcaption><span style=\"color:#505050\" class=\"has-inline-color\">Fig. 2: Forking-sequences<\/span><\/figcaption><\/figure>\n<p>This overlapping grid construction means forking-sequences fashions could be <strong>ensembled free of charge<\/strong> (or almost so) at inference time when it comes to saving encoder computation in contrast with window-sampling, which requires a number of impartial mannequin ahead passes. Given forecasts outputs through forking-sequences, we simply common (or in any other case mix) the totally different FCD-level predictions for a similar goal date portrayed because the diagonal band in Fig. 3a:<\/p>\n<p>[<br \/>\nbegin{equation}<br \/>\nwidetilde{mathbf{Y}}_{t,h} = frac{1}{H}sum_{k=0}^{H} widehat{mathbf{Y}}_{t-k, h+k} qquad text{for } tgeq H.<br \/>\nlabel{eq:forking_sequences_ensemble}<br \/>\nend{equation}<br \/>\n]<\/p>\n<div class=\"wp-block-columns\">\n<div class=\"wp-block-column\">\n<figure class=\"wp-block-image size-large\"><img decoding=\"async\" loading=\"lazy\" width=\"1024\" height=\"868\" src=\"https:\/\/blog.ml.cmu.edu\/wp-content\/uploads\/2026\/07\/available_forecast_page-0001-1024x868.jpg\" alt=\"\" class=\"wp-image-22597\" srcset=\"https:\/\/blog.ml.cmu.edu\/wp-content\/uploads\/2026\/07\/available_forecast_page-0001-1024x868.jpg 1024w, https:\/\/blog.ml.cmu.edu\/wp-content\/uploads\/2026\/07\/available_forecast_page-0001-300x254.jpg 300w, https:\/\/blog.ml.cmu.edu\/wp-content\/uploads\/2026\/07\/available_forecast_page-0001-1536x1302.jpg 1536w, https:\/\/blog.ml.cmu.edu\/wp-content\/uploads\/2026\/07\/available_forecast_page-0001-2048x1736.jpg 2048w, https:\/\/blog.ml.cmu.edu\/wp-content\/uploads\/2026\/07\/available_forecast_page-0001-737x625.jpg 737w, https:\/\/blog.ml.cmu.edu\/wp-content\/uploads\/2026\/07\/available_forecast_page-0001-271x230.jpg 271w, https:\/\/blog.ml.cmu.edu\/wp-content\/uploads\/2026\/07\/available_forecast_page-0001-260x220.jpg 260w, https:\/\/blog.ml.cmu.edu\/wp-content\/uploads\/2026\/07\/available_forecast_page-0001-80x68.jpg 80w, https:\/\/blog.ml.cmu.edu\/wp-content\/uploads\/2026\/07\/available_forecast_page-0001-300x254@2x.jpg 600w\" sizes=\"auto, (max-width: 1024px) 100vw, 1024px\"\/><figcaption><span style=\"color:#505050\" class=\"has-inline-color\">(a) Forking-sequences forecast ensemble<\/span><\/figcaption><\/figure>\n<\/div>\n<div class=\"wp-block-column\">\n<figure class=\"wp-block-image size-large\"><img decoding=\"async\" loading=\"lazy\" width=\"1024\" height=\"878\" src=\"https:\/\/blog.ml.cmu.edu\/wp-content\/uploads\/2026\/07\/forecast_variance_page-0001-1024x878.jpg\" alt=\"\" class=\"wp-image-22598\" srcset=\"https:\/\/blog.ml.cmu.edu\/wp-content\/uploads\/2026\/07\/forecast_variance_page-0001-1024x878.jpg 1024w, https:\/\/blog.ml.cmu.edu\/wp-content\/uploads\/2026\/07\/forecast_variance_page-0001-300x257.jpg 300w, https:\/\/blog.ml.cmu.edu\/wp-content\/uploads\/2026\/07\/forecast_variance_page-0001-1536x1317.jpg 1536w, https:\/\/blog.ml.cmu.edu\/wp-content\/uploads\/2026\/07\/forecast_variance_page-0001-2048x1755.jpg 2048w, https:\/\/blog.ml.cmu.edu\/wp-content\/uploads\/2026\/07\/forecast_variance_page-0001-729x625.jpg 729w, https:\/\/blog.ml.cmu.edu\/wp-content\/uploads\/2026\/07\/forecast_variance_page-0001-268x230.jpg 268w, https:\/\/blog.ml.cmu.edu\/wp-content\/uploads\/2026\/07\/forecast_variance_page-0001-257x220.jpg 257w, https:\/\/blog.ml.cmu.edu\/wp-content\/uploads\/2026\/07\/forecast_variance_page-0001-80x69.jpg 80w, https:\/\/blog.ml.cmu.edu\/wp-content\/uploads\/2026\/07\/forecast_variance_page-0001-300x257@2x.jpg 600w\" sizes=\"auto, (max-width: 1024px) 100vw, 1024px\"\/><figcaption><span style=\"color:#505050\" class=\"has-inline-color\">(b) Forecast volatility discount<\/span><\/figcaption><\/figure>\n<\/div>\n<\/div>\n<p class=\"has-small-font-size\"><span style=\"color:#505050\" class=\"has-inline-color\">Fig. 3: We adapt forking-sequences throughout inference to ensemble a number of forecasts of the identical future date by computing a operate (ex., shifting common) throughout predictions generated from earlier FCDs. b) Forking-sequences ensembling reduces forecast volatility, lowering the estimators variance with a linear convergence charge analogous to the weak legislation of enormous numbers.<\/span><\/p>\n<p>Though it&#8217;s tempting to anticipate a variance-reduction habits much like the outcomes of Theorem 1, you will need to acknowledge that forecast variance naturally will increase the additional a forecast is from its corresponding commentary. Consequently, there&#8217;s an inherent restrict to how a lot ensembling can cut back volatility: older forecast revisions carry considerably greater uncertainty, whereas more moderen revisions are each extra correct and fewer variable. This makes it fascinating for an ensemble to position higher weight on newer forecasts moderately than treating all revisions equally.<\/p>\n<h2>New Forecast Volatility Metrics<\/h2>\n<p>We introduce <strong>scaled Forecast proportion Change (sFPC)<\/strong> to measure the relative change in predicted quantiles throughout consecutive forecast creation dates, offering a quantitative view of temporal volatility or forecast revision charges. Impressed by the sMAPE metric, sFPC makes use of a symmetric denominator, based mostly on each present and former forecasts, to mitigate problems with numerical instability [1]. This design ensures robustness when coping with small predicted values and avoids the division-by-zero issues frequent in conventional percentage-based metrics.<\/p>\n<p>[<br \/>\nmathrm{sFPC}^{(q)}left(hat{mathbf{y}}^{(q)}_{[b][t][h]}proper)<br \/>\n= frac{200}{B occasions T occasions H} sum_{b,t,h} frac{|hat{y}^{(q)}_{b,t+1,h}-hat{y}^{(q)}_{b,t,h+1}|}{|hat{y}^{(q)}_{b,t+1,h}| + |hat{y}^{(q)}_{b,t,h+1}|} .<br \/>\n]<\/p>\n<p>Computing sFPC between consecutive forecasts treats <em>all<\/em> revisions as equally undesirable, even ones that clearly enhance accuracy. To deal with this, we additionally introduce <strong>scaled<\/strong> <strong>Extra Volatility (sEV)<\/strong>, a metric for probabilistic forecasts that solely penalizes revisions that transfer a forecast <em>away<\/em> from the reality, or that overshoot it. sEV is designed to reward accuracy-improving forecast revisions whereas distinguishing them from dangerous volatility. sEV is outlined as:<\/p>\n<p>[<br \/>\nmathrm{sEV}left(mathbf{y}_{[b][t][h]}, hat{mathbf{y}}_{[b][t][h]}proper) = frac{sum_{b,t,h} mathrm{EV}(y_{b,t,h}, mathbf{hat{y}}_{b,t,h+1}, mathbf{hat{y}}_{b,t+1,h})}{sum_{b,t,h} |y_{b,t,h}|}, quad textual content{the place}<br \/>\n]<\/p>\n<p>[<br \/>\nmathrm{EV}(y,;mathbf{hat{y}}_1,; mathbf{hat{y}}_2) = mathrm{QL}(mathbf{hat{y}}_2,mathbf{hat{y}}_1) &#8211; (mathrm{QL}(y,mathbf{hat{y}}_1)-mathrm{QL}(y,mathbf{hat{y}}_2)), quad text{and}<br \/>\n]<\/p>\n<p>[<br \/>\n    mathrm{QL}_q(y, hat{y}^{(q)}) = q(y-hat{y}^{(q)})_+ + (1-q)(hat{y}^{(q)}-y)_+  .<br \/>\n]<\/p>\n<p>EV has three helpful properties, confirmed formally within the paper:<\/p>\n<ul>\n<li><strong>Zero penalty for enhancing revisions<\/strong>, proven in Fig. 4a: if a revision strikes proportionally nearer to the bottom fact, touchdown on the direct path between the reality and the prior forecast, EV = 0.<\/li>\n<li><strong>Most penalty for deteriorating revisions<\/strong>, proven in Fig. 4b: if a revision strikes the forecast farther from the reality, with the previous forecast sitting between the reality and the brand new one, EV equals the total accuracy degradation, the distinction in quantile loss between the brand new forecast and the previous one.<\/li>\n<li><strong>Overshoot penalty<\/strong>, proven in Fig. 4c: if a revision strikes in the precise route however overshoots, with the reality touchdown between the previous and new forecast, EV penalizes solely the brand new forecast&#8217;s quantile loss towards the reality.<\/li>\n<\/ul>\n<div class=\"wp-block-columns\">\n<div class=\"wp-block-column\">\n<figure class=\"wp-block-image size-large\"><img decoding=\"async\" loading=\"lazy\" width=\"1024\" height=\"640\" src=\"https:\/\/blog.ml.cmu.edu\/wp-content\/uploads\/2026\/07\/zero-penalty__improving_revision_page-0001-1024x640.jpg\" alt=\"\" class=\"wp-image-22593\" srcset=\"https:\/\/blog.ml.cmu.edu\/wp-content\/uploads\/2026\/07\/zero-penalty__improving_revision_page-0001-1024x640.jpg 1024w, https:\/\/blog.ml.cmu.edu\/wp-content\/uploads\/2026\/07\/zero-penalty__improving_revision_page-0001-300x188.jpg 300w, https:\/\/blog.ml.cmu.edu\/wp-content\/uploads\/2026\/07\/zero-penalty__improving_revision_page-0001-1536x960.jpg 1536w, https:\/\/blog.ml.cmu.edu\/wp-content\/uploads\/2026\/07\/zero-penalty__improving_revision_page-0001-2048x1280.jpg 2048w, https:\/\/blog.ml.cmu.edu\/wp-content\/uploads\/2026\/07\/zero-penalty__improving_revision_page-0001-970x606.jpg 970w, https:\/\/blog.ml.cmu.edu\/wp-content\/uploads\/2026\/07\/zero-penalty__improving_revision_page-0001-320x200.jpg 320w, https:\/\/blog.ml.cmu.edu\/wp-content\/uploads\/2026\/07\/zero-penalty__improving_revision_page-0001-80x50.jpg 80w, https:\/\/blog.ml.cmu.edu\/wp-content\/uploads\/2026\/07\/zero-penalty__improving_revision_page-0001-300x188@2x.jpg 600w\" sizes=\"auto, (max-width: 1024px) 100vw, 1024px\"\/><figcaption><span style=\"color:#505050\" class=\"has-inline-color\">(a) Bettering revision<\/span><\/figcaption><\/figure>\n<\/div>\n<div class=\"wp-block-column\">\n<figure class=\"wp-block-image size-large\"><img decoding=\"async\" loading=\"lazy\" width=\"1024\" height=\"640\" src=\"https:\/\/blog.ml.cmu.edu\/wp-content\/uploads\/2026\/07\/maximum-penalty__degrading_revision_page-0001-1024x640.jpg\" alt=\"\" class=\"wp-image-22595\" srcset=\"https:\/\/blog.ml.cmu.edu\/wp-content\/uploads\/2026\/07\/maximum-penalty__degrading_revision_page-0001-1024x640.jpg 1024w, https:\/\/blog.ml.cmu.edu\/wp-content\/uploads\/2026\/07\/maximum-penalty__degrading_revision_page-0001-300x188.jpg 300w, https:\/\/blog.ml.cmu.edu\/wp-content\/uploads\/2026\/07\/maximum-penalty__degrading_revision_page-0001-1536x960.jpg 1536w, https:\/\/blog.ml.cmu.edu\/wp-content\/uploads\/2026\/07\/maximum-penalty__degrading_revision_page-0001-2048x1280.jpg 2048w, https:\/\/blog.ml.cmu.edu\/wp-content\/uploads\/2026\/07\/maximum-penalty__degrading_revision_page-0001-970x606.jpg 970w, https:\/\/blog.ml.cmu.edu\/wp-content\/uploads\/2026\/07\/maximum-penalty__degrading_revision_page-0001-320x200.jpg 320w, https:\/\/blog.ml.cmu.edu\/wp-content\/uploads\/2026\/07\/maximum-penalty__degrading_revision_page-0001-80x50.jpg 80w, https:\/\/blog.ml.cmu.edu\/wp-content\/uploads\/2026\/07\/maximum-penalty__degrading_revision_page-0001-300x188@2x.jpg 600w\" sizes=\"auto, (max-width: 1024px) 100vw, 1024px\"\/><figcaption><span style=\"color:#505050\" class=\"has-inline-color\">(b) Deteriorating revision<\/span><\/figcaption><\/figure>\n<\/div>\n<div class=\"wp-block-column\">\n<figure class=\"wp-block-image size-large\"><img decoding=\"async\" loading=\"lazy\" width=\"1024\" height=\"640\" src=\"https:\/\/blog.ml.cmu.edu\/wp-content\/uploads\/2026\/07\/overshoot-revision_penalty_page-0001-1024x640.jpg\" alt=\"\" class=\"wp-image-22594\" srcset=\"https:\/\/blog.ml.cmu.edu\/wp-content\/uploads\/2026\/07\/overshoot-revision_penalty_page-0001-1024x640.jpg 1024w, https:\/\/blog.ml.cmu.edu\/wp-content\/uploads\/2026\/07\/overshoot-revision_penalty_page-0001-300x188.jpg 300w, https:\/\/blog.ml.cmu.edu\/wp-content\/uploads\/2026\/07\/overshoot-revision_penalty_page-0001-1536x960.jpg 1536w, https:\/\/blog.ml.cmu.edu\/wp-content\/uploads\/2026\/07\/overshoot-revision_penalty_page-0001-2048x1280.jpg 2048w, https:\/\/blog.ml.cmu.edu\/wp-content\/uploads\/2026\/07\/overshoot-revision_penalty_page-0001-970x606.jpg 970w, https:\/\/blog.ml.cmu.edu\/wp-content\/uploads\/2026\/07\/overshoot-revision_penalty_page-0001-320x200.jpg 320w, https:\/\/blog.ml.cmu.edu\/wp-content\/uploads\/2026\/07\/overshoot-revision_penalty_page-0001-80x50.jpg 80w, https:\/\/blog.ml.cmu.edu\/wp-content\/uploads\/2026\/07\/overshoot-revision_penalty_page-0001-300x188@2x.jpg 600w\" sizes=\"auto, (max-width: 1024px) 100vw, 1024px\"\/><figcaption><span style=\"color:#505050\" class=\"has-inline-color\">(c) Overshooting revision<\/span><\/figcaption><\/figure>\n<\/div>\n<\/div>\n<p class=\"has-small-font-size\"><span style=\"color:#505050\" class=\"has-inline-color\">Fig. 4: Instance penalty habits of the Extra Volatility (EV) metric. EV distinguishes accuracy-improving revisions from accuracy-degrading ones, assigning no penalty when revisions enhance accuracy, whereas asymmetrically penalizing each deteriorating and overshooting revisions in response to their impression on accuracy.<\/span><\/p>\n<p>One necessary distinction: sFPC could be computed at prediction time for real-time monitoring, because it would not require floor fact. sEV, against this, depends upon the ground-truth worth, so it could solely be utilized retroactively to evaluate forecast volatility.<\/p>\n<h2>Empirical Outcomes: Volatility Discount With out Sacrificing Accuracy<\/h2>\n<p>The core empirical declare: for forking-sequences fashions, making use of exponential-smoothing ensembling at inference (\u03b1 = 0.9) reduces forecast volatility (sEV) considerably <strong>whereas sustaining forecast accuracy.<\/strong><\/p>\n<p>We present that for forking-sequences fashions, forecast ensembling throughout inference can cut back forecast volatility in comparison with forecasts with out ensembling for all encoders. Particularly, making use of exponential smoothing at inference to fashions skilled with forking-sequences yields median proportion enhancements in sEV throughout datasets of 13.2%, 13.0%, 10.9%, 10.2%, and 11.2% for RNN, LSTM, CNN, Transformer, and StateSpace-based architectures, respectively, whereas sustaining forecast accuracy (<strong>lower than 0.1% degradation in sCRPS<\/strong> as proven in Fig. 5).<\/p>\n<div class=\"wp-block-columns\">\n<div class=\"wp-block-column\">\n<figure class=\"wp-block-image size-large\"><img decoding=\"async\" loading=\"lazy\" width=\"1024\" height=\"637\" src=\"https:\/\/blog.ml.cmu.edu\/wp-content\/uploads\/2026\/07\/scrps_boxplot_exponential_smoothing_alpha9_ensemble_page-0001-1024x637.jpg\" alt=\"\" class=\"wp-image-22601\" srcset=\"https:\/\/blog.ml.cmu.edu\/wp-content\/uploads\/2026\/07\/scrps_boxplot_exponential_smoothing_alpha9_ensemble_page-0001-1024x637.jpg 1024w, https:\/\/blog.ml.cmu.edu\/wp-content\/uploads\/2026\/07\/scrps_boxplot_exponential_smoothing_alpha9_ensemble_page-0001-300x186.jpg 300w, https:\/\/blog.ml.cmu.edu\/wp-content\/uploads\/2026\/07\/scrps_boxplot_exponential_smoothing_alpha9_ensemble_page-0001-1536x955.jpg 1536w, https:\/\/blog.ml.cmu.edu\/wp-content\/uploads\/2026\/07\/scrps_boxplot_exponential_smoothing_alpha9_ensemble_page-0001-2048x1273.jpg 2048w, https:\/\/blog.ml.cmu.edu\/wp-content\/uploads\/2026\/07\/scrps_boxplot_exponential_smoothing_alpha9_ensemble_page-0001-970x603.jpg 970w, https:\/\/blog.ml.cmu.edu\/wp-content\/uploads\/2026\/07\/scrps_boxplot_exponential_smoothing_alpha9_ensemble_page-0001-320x199.jpg 320w, https:\/\/blog.ml.cmu.edu\/wp-content\/uploads\/2026\/07\/scrps_boxplot_exponential_smoothing_alpha9_ensemble_page-0001-80x50.jpg 80w, https:\/\/blog.ml.cmu.edu\/wp-content\/uploads\/2026\/07\/scrps_boxplot_exponential_smoothing_alpha9_ensemble_page-0001-300x186@2x.jpg 600w\" sizes=\"auto, (max-width: 1024px) 100vw, 1024px\"\/><figcaption><span style=\"color:#505050\" class=\"has-inline-color\">(a) sCRPS<\/span><\/figcaption><\/figure>\n<\/div>\n<div class=\"wp-block-column\">\n<figure class=\"wp-block-image size-large\"><img decoding=\"async\" loading=\"lazy\" width=\"1024\" height=\"637\" src=\"https:\/\/blog.ml.cmu.edu\/wp-content\/uploads\/2026\/07\/sev_boxplot_exponential_smoothing_alpha9_ensemble_page-0001-1024x637.jpg\" alt=\"\" class=\"wp-image-22602\" srcset=\"https:\/\/blog.ml.cmu.edu\/wp-content\/uploads\/2026\/07\/sev_boxplot_exponential_smoothing_alpha9_ensemble_page-0001-1024x637.jpg 1024w, https:\/\/blog.ml.cmu.edu\/wp-content\/uploads\/2026\/07\/sev_boxplot_exponential_smoothing_alpha9_ensemble_page-0001-300x186.jpg 300w, https:\/\/blog.ml.cmu.edu\/wp-content\/uploads\/2026\/07\/sev_boxplot_exponential_smoothing_alpha9_ensemble_page-0001-1536x955.jpg 1536w, https:\/\/blog.ml.cmu.edu\/wp-content\/uploads\/2026\/07\/sev_boxplot_exponential_smoothing_alpha9_ensemble_page-0001-2048x1273.jpg 2048w, https:\/\/blog.ml.cmu.edu\/wp-content\/uploads\/2026\/07\/sev_boxplot_exponential_smoothing_alpha9_ensemble_page-0001-970x603.jpg 970w, https:\/\/blog.ml.cmu.edu\/wp-content\/uploads\/2026\/07\/sev_boxplot_exponential_smoothing_alpha9_ensemble_page-0001-320x199.jpg 320w, https:\/\/blog.ml.cmu.edu\/wp-content\/uploads\/2026\/07\/sev_boxplot_exponential_smoothing_alpha9_ensemble_page-0001-80x50.jpg 80w, https:\/\/blog.ml.cmu.edu\/wp-content\/uploads\/2026\/07\/sev_boxplot_exponential_smoothing_alpha9_ensemble_page-0001-300x186@2x.jpg 600w\" sizes=\"auto, (max-width: 1024px) 100vw, 1024px\"\/><figcaption><span style=\"color:#505050\" class=\"has-inline-color\">(b) sEV<\/span><\/figcaption><\/figure>\n<\/div>\n<div class=\"wp-block-column\">\n<figure class=\"wp-block-image size-large\"><img decoding=\"async\" loading=\"lazy\" width=\"1024\" height=\"637\" src=\"https:\/\/blog.ml.cmu.edu\/wp-content\/uploads\/2026\/07\/sqpc_boxplot_exponential_smoothing_alpha9_ensemble_page-0001-1024x637.jpg\" alt=\"\" class=\"wp-image-22603\" srcset=\"https:\/\/blog.ml.cmu.edu\/wp-content\/uploads\/2026\/07\/sqpc_boxplot_exponential_smoothing_alpha9_ensemble_page-0001-1024x637.jpg 1024w, https:\/\/blog.ml.cmu.edu\/wp-content\/uploads\/2026\/07\/sqpc_boxplot_exponential_smoothing_alpha9_ensemble_page-0001-300x186.jpg 300w, https:\/\/blog.ml.cmu.edu\/wp-content\/uploads\/2026\/07\/sqpc_boxplot_exponential_smoothing_alpha9_ensemble_page-0001-1536x955.jpg 1536w, https:\/\/blog.ml.cmu.edu\/wp-content\/uploads\/2026\/07\/sqpc_boxplot_exponential_smoothing_alpha9_ensemble_page-0001-2048x1273.jpg 2048w, https:\/\/blog.ml.cmu.edu\/wp-content\/uploads\/2026\/07\/sqpc_boxplot_exponential_smoothing_alpha9_ensemble_page-0001-970x603.jpg 970w, https:\/\/blog.ml.cmu.edu\/wp-content\/uploads\/2026\/07\/sqpc_boxplot_exponential_smoothing_alpha9_ensemble_page-0001-320x199.jpg 320w, https:\/\/blog.ml.cmu.edu\/wp-content\/uploads\/2026\/07\/sqpc_boxplot_exponential_smoothing_alpha9_ensemble_page-0001-80x50.jpg 80w, https:\/\/blog.ml.cmu.edu\/wp-content\/uploads\/2026\/07\/sqpc_boxplot_exponential_smoothing_alpha9_ensemble_page-0001-300x186@2x.jpg 600w\" sizes=\"auto, (max-width: 1024px) 100vw, 1024px\"\/><figcaption><span style=\"color:#505050\" class=\"has-inline-color\">(c) sFPC<\/span><\/figcaption><\/figure>\n<\/div>\n<\/div>\n<p class=\"has-small-font-size\"><span style=\"color:#505050\" class=\"has-inline-color\">Fig. 5: Distribution of proportion enchancment in (a) sCRPS, (b) sEV, and (c) sFPC metrics throughout datasets for various encoder varieties with forking-sequences forecast ensembling in contrast with no ensembling. Every dataset&#8217;s metric is averaged over 5 random seed runs. Proportion enchancment higher than zero signifies forecast ensembling achieves decrease forecast error or volatility.<\/span><\/p>\n<p>We embody an ablation research throughout totally different ensembling methods (shifting common, shifting median, cumulative common, exponential smoothing at \u03b1 = 0.1\/0.5\/0.9), and discover that <strong>exponential smoothing with excessive \u03b1 (0.9) offers the most effective trade-off<\/strong>; it weights near-term (extra correct) forecasts extra closely, minimizing the accuracy value of smoothing out volatility. Decrease \u03b1 values cut back volatility additional however at a better value to accuracy.<\/p>\n<h2>A Bonus: Zero-Shot Volatility Reductions for Pretrained Basis Fashions<\/h2>\n<p>Forecast ensembling profit <strong>is not restricted to fashions particularly skilled with forking-sequences<\/strong>. We are able to apply forecast ensembling to pretrained fashions initially skilled with window-sampling by accumulating forecast revision outputs. We reveal this with pretrained Time Collection Basis Fashions (TSFMs), together with Chronos-2, Toto 2.0, TimesFM, and pretrained PatchTST and NBEATS, in a zero-shot setting.<\/p>\n<div class=\"wp-block-columns\">\n<div class=\"wp-block-column\">\n<figure class=\"wp-block-image size-large is-resized\"><img decoding=\"async\" loading=\"lazy\" src=\"https:\/\/blog.ml.cmu.edu\/wp-content\/uploads\/2026\/07\/scrps_boxplot_exponential_smoothing9_ensemble_zeroshot_page-0001-1024x724.jpg\" alt=\"\" class=\"wp-image-22590\" width=\"172\" height=\"121\" srcset=\"https:\/\/blog.ml.cmu.edu\/wp-content\/uploads\/2026\/07\/scrps_boxplot_exponential_smoothing9_ensemble_zeroshot_page-0001-1024x724.jpg 1024w, https:\/\/blog.ml.cmu.edu\/wp-content\/uploads\/2026\/07\/scrps_boxplot_exponential_smoothing9_ensemble_zeroshot_page-0001-300x212.jpg 300w, https:\/\/blog.ml.cmu.edu\/wp-content\/uploads\/2026\/07\/scrps_boxplot_exponential_smoothing9_ensemble_zeroshot_page-0001-1536x1086.jpg 1536w, https:\/\/blog.ml.cmu.edu\/wp-content\/uploads\/2026\/07\/scrps_boxplot_exponential_smoothing9_ensemble_zeroshot_page-0001-2048x1448.jpg 2048w, https:\/\/blog.ml.cmu.edu\/wp-content\/uploads\/2026\/07\/scrps_boxplot_exponential_smoothing9_ensemble_zeroshot_page-0001-884x625.jpg 884w, https:\/\/blog.ml.cmu.edu\/wp-content\/uploads\/2026\/07\/scrps_boxplot_exponential_smoothing9_ensemble_zeroshot_page-0001-320x226.jpg 320w, https:\/\/blog.ml.cmu.edu\/wp-content\/uploads\/2026\/07\/scrps_boxplot_exponential_smoothing9_ensemble_zeroshot_page-0001-80x57.jpg 80w, https:\/\/blog.ml.cmu.edu\/wp-content\/uploads\/2026\/07\/scrps_boxplot_exponential_smoothing9_ensemble_zeroshot_page-0001-300x212@2x.jpg 600w\" sizes=\"auto, (max-width: 172px) 100vw, 172px\"\/><figcaption><span style=\"color:#505050\" class=\"has-inline-color\">(a) sCRPS<\/span><\/figcaption><\/figure>\n<\/div>\n<div class=\"wp-block-column\">\n<figure class=\"wp-block-image size-large is-resized\"><img decoding=\"async\" loading=\"lazy\" src=\"https:\/\/blog.ml.cmu.edu\/wp-content\/uploads\/2026\/07\/sev_boxplot_exponential_smoothing9_ensemble_zeroshot_page-0001-1024x724.jpg\" alt=\"\" class=\"wp-image-22591\" width=\"172\" height=\"121\" srcset=\"https:\/\/blog.ml.cmu.edu\/wp-content\/uploads\/2026\/07\/sev_boxplot_exponential_smoothing9_ensemble_zeroshot_page-0001-1024x724.jpg 1024w, https:\/\/blog.ml.cmu.edu\/wp-content\/uploads\/2026\/07\/sev_boxplot_exponential_smoothing9_ensemble_zeroshot_page-0001-300x212.jpg 300w, https:\/\/blog.ml.cmu.edu\/wp-content\/uploads\/2026\/07\/sev_boxplot_exponential_smoothing9_ensemble_zeroshot_page-0001-1536x1086.jpg 1536w, https:\/\/blog.ml.cmu.edu\/wp-content\/uploads\/2026\/07\/sev_boxplot_exponential_smoothing9_ensemble_zeroshot_page-0001-2048x1448.jpg 2048w, https:\/\/blog.ml.cmu.edu\/wp-content\/uploads\/2026\/07\/sev_boxplot_exponential_smoothing9_ensemble_zeroshot_page-0001-884x625.jpg 884w, https:\/\/blog.ml.cmu.edu\/wp-content\/uploads\/2026\/07\/sev_boxplot_exponential_smoothing9_ensemble_zeroshot_page-0001-320x226.jpg 320w, https:\/\/blog.ml.cmu.edu\/wp-content\/uploads\/2026\/07\/sev_boxplot_exponential_smoothing9_ensemble_zeroshot_page-0001-80x57.jpg 80w, https:\/\/blog.ml.cmu.edu\/wp-content\/uploads\/2026\/07\/sev_boxplot_exponential_smoothing9_ensemble_zeroshot_page-0001-300x212@2x.jpg 600w\" sizes=\"auto, (max-width: 172px) 100vw, 172px\"\/><figcaption><span style=\"color:#505050\" class=\"has-inline-color\">(b) sEV<\/span><\/figcaption><\/figure>\n<\/div>\n<div class=\"wp-block-column\">\n<figure class=\"wp-block-image size-large\"><img decoding=\"async\" loading=\"lazy\" width=\"1024\" height=\"724\" src=\"https:\/\/blog.ml.cmu.edu\/wp-content\/uploads\/2026\/07\/fpc_boxplot_exponential_smoothing9_ensemble_zeroshot_page-0001-1024x724.jpg\" alt=\"\" class=\"wp-image-22592\" srcset=\"https:\/\/blog.ml.cmu.edu\/wp-content\/uploads\/2026\/07\/fpc_boxplot_exponential_smoothing9_ensemble_zeroshot_page-0001-1024x724.jpg 1024w, https:\/\/blog.ml.cmu.edu\/wp-content\/uploads\/2026\/07\/fpc_boxplot_exponential_smoothing9_ensemble_zeroshot_page-0001-300x212.jpg 300w, https:\/\/blog.ml.cmu.edu\/wp-content\/uploads\/2026\/07\/fpc_boxplot_exponential_smoothing9_ensemble_zeroshot_page-0001-1536x1086.jpg 1536w, https:\/\/blog.ml.cmu.edu\/wp-content\/uploads\/2026\/07\/fpc_boxplot_exponential_smoothing9_ensemble_zeroshot_page-0001-2048x1448.jpg 2048w, https:\/\/blog.ml.cmu.edu\/wp-content\/uploads\/2026\/07\/fpc_boxplot_exponential_smoothing9_ensemble_zeroshot_page-0001-884x625.jpg 884w, https:\/\/blog.ml.cmu.edu\/wp-content\/uploads\/2026\/07\/fpc_boxplot_exponential_smoothing9_ensemble_zeroshot_page-0001-320x226.jpg 320w, https:\/\/blog.ml.cmu.edu\/wp-content\/uploads\/2026\/07\/fpc_boxplot_exponential_smoothing9_ensemble_zeroshot_page-0001-80x57.jpg 80w, https:\/\/blog.ml.cmu.edu\/wp-content\/uploads\/2026\/07\/fpc_boxplot_exponential_smoothing9_ensemble_zeroshot_page-0001-300x212@2x.jpg 600w\" sizes=\"auto, (max-width: 1024px) 100vw, 1024px\"\/><figcaption><span style=\"color:#505050\" class=\"has-inline-color\">(c) sFPC<\/span><\/figcaption><\/figure>\n<\/div>\n<\/div>\n<p class=\"has-small-font-size\"><span style=\"color:#505050\" class=\"has-inline-color\">Fig. 6:  Distribution of proportion enchancment in (a) sCRPS, (b) sEV, and (c) sFPC metrics throughout datasets for various encoder varieties with forking-sequences forecast ensembling in contrast with no ensembling.<\/span> <span style=\"color:#505050\" class=\"has-inline-color\">Proportion enchancment higher than zero signifies forecast ensembling achieves decrease forecast error or volatility. Forecast ensembling can considerably cut back forecast volatility (sEV, sFPC) whereas sustaining forecast accuracy (sCRPS), demonstrating its utility as a general-purpose inference method for fashions skilled with both forking-sequences or window-sampling.<\/span><\/p>\n<p>Throughout the M-series benchmark, this easy method achieved a <strong>median ~10% discount in forecast volatility<\/strong>, with <strong>lower than 0.1% degradation in accuracy (sCRPS)<\/strong>. In different phrases: forecast ensembling through forking-sequences-style aggregation is a general-purpose, nearly-free method that can be utilized in forecasting pipelines no matter whether or not the underlying mannequin was initially skilled with forking-sequences.<\/p>\n<h2>\ud83d\udd11 Takeaways<\/h2>\n<div style=\"border: 1px solid #e0e0e0; border-radius: 10px; padding: 24px 28px; margin: 28px 0; font-family: -apple-system, BlinkMacSystemFont, 'Segoe UI', Helvetica, Arial, sans-serif; box-shadow: 0 1px 3px rgba(0,0,0,0.06);\">\n<p>\n    <span style=\"flex-shrink: 0; width: 26px; height: 26px; border-radius: 50%; background: #2f855a; color: #fff; font-size: 13px; font-weight: 700; display: flex; align-items: center; justify-content: center; margin-right: 14px;\">1<\/span><br \/>\n    <span style=\"color: #2a2a2a; font-size: 16px; line-height: 1.6;\">Forking-sequences&#8217; grid construction naturally produces overlapping forecasts throughout FCDs, enabling near-free ensembling at inference time by reusing already-computed encoder outputs.<\/span>\n  <\/p>\n<p>\n    <span style=\"flex-shrink: 0; width: 26px; height: 26px; border-radius: 50%; background: #2f855a; color: #fff; font-size: 13px; font-weight: 700; display: flex; align-items: center; justify-content: center; margin-right: 14px;\">2<\/span><br \/>\n    <span style=\"color: #2a2a2a; font-size: 16px; line-height: 1.6;\">The brand new scaled Extra Volatility (sEV) metric distinguishes accuracy-improving revisions from dangerous ones \u2014 a significant enchancment over naive percentage-change volatility measures.<\/span>\n  <\/p>\n<p>\n    <span style=\"flex-shrink: 0; width: 26px; height: 26px; border-radius: 50%; background: #2f855a; color: #fff; font-size: 13px; font-weight: 700; display: flex; align-items: center; justify-content: center; margin-right: 14px;\">3<\/span><br \/>\n    <span style=\"color: #2a2a2a; font-size: 16px; line-height: 1.6;\">Ensembling forking-sequences forecasts through exponential smoothing cuts volatility by ~10\u201313% throughout encoder architectures with out sacrificing accuracy.<\/span>\n  <\/p>\n<p>\n    <span style=\"flex-shrink: 0; width: 26px; height: 26px; border-radius: 50%; background: #2f855a; color: #fff; font-size: 13px; font-weight: 700; display: flex; align-items: center; justify-content: center; margin-right: 14px;\">4<\/span><br \/>\n    <span style=\"color: #2a2a2a; font-size: 16px; line-height: 1.6;\">This profit extends to zero-shot use with pretrained basis fashions like Chronos-2, Toto 2.0, and TimesFM, attaining roughly 10% diminished forecast volatility with &lt;0.1% accuracy value.<\/span>\n  <\/p>\n<\/div>\n<p>We acknowledge that ensembling could be built-in throughout each coaching and inference with forking-sequences, and might be additional prolonged with learnable parameters as explored in [2]. We go away training-time ensembling integration to future work.<\/p>\n<p>Collectively, Components I and II purpose to construct broader consciousness of forking-sequences and promote its adoption as a default architectural possibility in open-source neural forecasting libraries and future analysis. This work additionally advocates for higher consciousness of volatility metrics as a complement to straightforward accuracy metrics, encouraging their routine adoption in forecasting analysis.<\/p>\n<hr class=\"wp-block-separator\"\/>\n<p><strong>References:<\/strong><br \/>[1] Rob J. Hyndman and Anne B. Koehler. One other take a look at measures of forecast accuracy. Worldwide Journal of Forecasting, 22(4):679 \u2013 688, 2006. ISSN 0169-2070.<\/p>\n<p>[2] Carson Eisenach, Yagna Patel, and Dhruv Madeka. MQTransformer: Multi-Horizon Forecasts with Context Dependent and Suggestions-Conscious Consideration. In Maria Florina Balcan and Marina Meila, editors, Submitted to Proceedings of the thirty eighth Worldwide Convention on Machine Studying. PMLR. Working Paper model out there at arXiv:2009.14799, 8 2021.<\/p>\n<\/p><\/div>\n\n","protected":false},"excerpt":{"rendered":"<p>Primarily based on: Potosnak, W., Wolff, M., Cao, M., Ma, R., Konstantinova, T., Efimov, D., Mahoney, M.W., Oreshkin, B., &amp; Olivares, Ok.G. &#8220;Forking-Sequences: Statistically and Computationally Environment friendly Multi-Horizon Forecasting with Decreased Volatility.&#8221; Transactions on Machine Studying Analysis, 2026. (Disclaimer: Code implementation not used within the paper; not affiliated with Amazon \u2014 offered as a [&hellip;]<\/p>\n","protected":false},"author":2,"featured_media":17689,"comment_status":"open","ping_status":"open","sticky":false,"template":"","format":"standard","meta":{"footnotes":""},"categories":[55],"tags":[110,10145,5187,10143,136,113,442,10144,668,7812,10146],"class_list":["post-17687","post","type-post","status-publish","format-standard","has-post-thumbnail","hentry","category-machine-learning","tag-blog","tag-ensembling","tag-forecast","tag-forkingsequences","tag-learning","tag-machine","tag-mlcmu","tag-multihorizon","tag-part","tag-reduced","tag-volatility"],"_links":{"self":[{"href":"https:\/\/techtrendfeed.com\/index.php?rest_route=\/wp\/v2\/posts\/17687","targetHints":{"allow":["GET"]}}],"collection":[{"href":"https:\/\/techtrendfeed.com\/index.php?rest_route=\/wp\/v2\/posts"}],"about":[{"href":"https:\/\/techtrendfeed.com\/index.php?rest_route=\/wp\/v2\/types\/post"}],"author":[{"embeddable":true,"href":"https:\/\/techtrendfeed.com\/index.php?rest_route=\/wp\/v2\/users\/2"}],"replies":[{"embeddable":true,"href":"https:\/\/techtrendfeed.com\/index.php?rest_route=%2Fwp%2Fv2%2Fcomments&post=17687"}],"version-history":[{"count":1,"href":"https:\/\/techtrendfeed.com\/index.php?rest_route=\/wp\/v2\/posts\/17687\/revisions"}],"predecessor-version":[{"id":17688,"href":"https:\/\/techtrendfeed.com\/index.php?rest_route=\/wp\/v2\/posts\/17687\/revisions\/17688"}],"wp:featuredmedia":[{"embeddable":true,"href":"https:\/\/techtrendfeed.com\/index.php?rest_route=\/wp\/v2\/media\/17689"}],"wp:attachment":[{"href":"https:\/\/techtrendfeed.com\/index.php?rest_route=%2Fwp%2Fv2%2Fmedia&parent=17687"}],"wp:term":[{"taxonomy":"category","embeddable":true,"href":"https:\/\/techtrendfeed.com\/index.php?rest_route=%2Fwp%2Fv2%2Fcategories&post=17687"},{"taxonomy":"post_tag","embeddable":true,"href":"https:\/\/techtrendfeed.com\/index.php?rest_route=%2Fwp%2Fv2%2Ftags&post=17687"}],"curies":[{"name":"wp","href":"https:\/\/api.w.org\/{rel}","templated":true}]}}<!-- This website is optimized by Airlift. Learn more: https://airlift.net. Template:. Learn more: https://airlift.net. Template: 69d9690a190636c2e0989534. Config Timestamp: 2026-04-10 21:18:02 UTC, Cached Timestamp: 2026-08-13 03:09:18 UTC -->