Waiting for enough evidence.
Experimental River Flow Forecasts for Pagosa Springs
Questionable hats. Serious data. Experimental forecasts.
In a top secret-ish corner of Pagosa Outside, we wired 25 years of river records, snowpack, weather forecasts, strange local theories, snowfield patterns and finger-in-the-wind hunches into one questionable contraption. It sparks, rattles and spits out educated guesses while Mother Nature keeps score—because apparently asking a groundhog to predict spring is tradition, but using actual data is mad science. Welcome to Pagosa Outside’s Black Hat Hydrology Lab.
This is experimental forecasting and should not be relied on for anything other than entertainment. Use the River Desk and official sources for current conditions and published forecasts.
What the Black Hat sees right now
This is the Lab’s current working answer—not one of the locked contest predictions. It updates as the river, headwater weather and seasonal evidence change.
Checking the river’s next move…
The Lab is comparing the newest river readings with forecast headwater weather and similar moments from previous water years.
Waiting for enough evidence.
Waiting for the newest observations.
The forecast feed is still loading.
The Lab is finding comparable seasons.
New observations, forecast weather or a meaningful change in the river’s direction can move the live estimate.
Run the Lab again after the next source update to see whether the evidence moved.
The current answer can move. The official contest guesses farther down the page cannot.
The official eight-prediction game
Eight locked questions, four quarters and one opponent who refuses to share her playbook.
PO BLACK HAT LAB vs MOTHER NATURE
Every October 1 starts a new water-year game. A prediction is locked on its posted issue date and cannot move after better information arrives. A hit scores one for the Black Hat Lab; a miss scores one for Mother Nature. Results are credited to the quarter in which they become final, so the two full-season counts remain pending until after Labor Day. This first season was recreated without peeking beyond each issue date—and every miss stays on the board.
Where the eight locked predictions stand
The predictions below were made on fixed issue dates. This tracker follows them to their finish lines without moving the original guesses.
- 1Snow sticksMISS
- 2Peak snowMISS
- 3First 400+HIT
- 4River strengthHIT
- 5Spring peakHIT
- 6Rafting daysPENDING
- 7Tubing opensMISS
- 8Tubing daysPENDING
Loading the active predictions…
The Lab is checking which locked questions are settled, still alive or waiting for their scoring date.
Waiting for the current evidence.
The next finish line will appear here.
Eight questions people actually want answered
When will snow stick? How much water will winter bank? When will rafting flows arrive, how big will runoff get and when will the river hand the keys to tubing? These are fun facts with real consequences: several also help Pagosa Outside think about staffing, equipment, trip mix and the timing of a short river season.
Gold is the locked Black Hat guess, blue is the 25-year historical average and green is what Mother Nature actually delivered. Each question gets its own model, finish line and scoring rule.
First persistent headwater snowpack
What we’re predicting: When winter will stop teasing us and leave a lasting snow reserve in the San Juan headwaters.
- Black Hat guess
- October 24, 2025
- Historical average
- October 24, 2025
- Actual result
- November 17, 2025
Result: Mother Nature waited 24 days longer than our guess.
Win condition: Land within 14 days of the first date the two-station average reaches one inch of stored water and remains above one-half inch for seven days.
Peak headwater snowpack
What we’re predicting: How much water winter will bank in mountain snow before melting begins—an early clue to runoff strength and season length.
- Black Hat guess
- 37.3 in · April 2, 2026
- Historical average
- 27.3 in · April 6, 2026
- Actual result
- 14.3 in · March 6, 2026
Result: We overshot the peak by 23 inches. The miss exposed a low-snow failure in the old tree model, so the next build gets a stricter warm-and-dry guardrail.
Win condition: Predict the peak amount within 20% or two inches, whichever gives the wider scoring window.
First 400+ cubic-feet-per-second day
What we’re predicting: When the river will first climb out of winter flow and reach Pagosa’s general rafting crossover.
- Black Hat guess
- March 29, 2026
- Historical average
- April 2, 2026
- Actual result
- March 20, 2026 · 421 CFS
Result: The river arrived nine days earlier than our guess—close enough to put one on the Black Hat side.
Win condition: Land within 10 days of the first calendar day whose average flow reaches at least 400 CFS; a brief spike does not count.
May–June rafting-season strength
What we’re predicting: Whether Pagosa is headed for a high-water year locals should not miss, a dependable average rafting year or a mild season when guests should get on the river early instead of waiting for warm June weather.
- Black Hat guess
- 376 CFS average
MILD RAFTING YEAR - Historical average
- About 1,100 CFS
AVERAGE RAFTING YEAR - Actual result
- 341 CFS average
MILD RAFTING YEAR
Result: Only 35 CFS high. The Lab correctly warned that this would be a mild rafting year—one to catch early rather than postponing until June warmed up.
Win condition: Predict the May 1–June 30 average within 30% or 100 CFS. Mild is below 75% of the historical average; high-water begins above 125%.
Spring peak flow
What we’re predicting: The date, Mountain Time and flow of the spring crest—the classic runoff question and valuable high-water planning context.
- Black Hat guess
- May 14, 2026 · 12:15 a.m.
984 CFS - Historical average
- May 22, 2026 · 12:15 a.m.
2,280 CFS - Actual result
- May 15, 2026 · 1:15 a.m.
900 CFS
Result: One day early, one hour early and 84 CFS high—a clean Black Hat win.
Win condition: Get the flow within 30% or 250 CFS and also place either the date within seven days or the time within two hours.
Rafting-flow days, May 1–Labor Day
What we’re predicting: How many days will reach the flow benchmarks behind the season’s Splash & Dash, Mesa Canyon and West Fork planning.
- Black Hat guess
- 45 at 300+
31 at 600+ · 11 at 800+ - Historical average
- 63 at 300+
46 at 600+ · 39 at 800+ - Through August 21, 2026
- 36 at 300+
8 at 600+ · 0 at 800+
Result: Still cooking. The 300-CFS count remains within reach; the 600- and 800-CFS guesses are already far above the season so far.
Win condition: Finish within 10 days on at least two of the three counts after Labor Day.
Tubing Season Opening Day
What we’re predicting: When post-peak runoff will hand the river from spring rafting to warm-weather tubing.
- Black Hat guess
- June 14, 2026
- Historical average
- June 24, 2026
- Actual result
- June 4, 2026
Result: Tubing conditions arrived 10 days earlier than our guess, putting this point on Mother Nature’s side.
Win condition: Land within seven days of the first complete post-peak 24-hour period with every available gauge reading from 30 through 399 CFS.
Tubing-range days, May 1–Labor Day
What we’re predicting: How many recreation-season days will land in Pagosa’s broad tubing range—an estimate of the season’s usable tubing supply.
- Black Hat guess
- 77 days
- Historical average
- 66 days
- Through August 21, 2026
- 44 days
25 too high · 44 too low
Result: Still cooking. Low water—not rafting water—has removed most of the missing tubing days.
Win condition: Finish within 10 days of the total number of May 1–Labor Day daily averages from 30 through 399 CFS.
How the river may behave from here
Three fixed outlooks, one continuous hydrograph and the same operating-threshold probabilities viewed at several distances.
Calculating the 30-day experiment.
Calculating the 60-day experiment.
Calculating the 90-day experiment.
Possible flow over the next 90 days
| Forecast date | Center estimate | Experimental range | 30–400 CFS | 300+ CFS | 600+ CFS | 800+ CFS |
|---|
WHAT IS UNDER THE BLACK HAT?
The live 30-, 60- and 90-day outlook uses the current Pagosa gauge, recent flow movement, Upper San Juan and Wolf Creek Summit snowpack, headwater temperature, accumulated precipitation and seasonal timing. Archived NOAA climate outlooks were tested here too, but they did not improve these daily-flow, hydrograph or operating-threshold forecasts enough to earn model weight.
From October through February the Lab switches from a river-flow hydrograph to experimental snowpack development. The Right Now in the Lab briefing changes its question with the river year: lasting snow, peak snowpack, rafting arrival, spring peak, the tubing transition, monsoon recovery and fall bonus flow. It always shows the current working answer, the evidence behind it and what could change it. The separate Season Game Tracker follows the eight locked predictions without rewriting their original guesses.
The spring peak target is the highest U.S. Geological Survey instantaneous reading at the Pagosa gauge from April 1 through July 31. The Black Hat guess reports one date, one Mountain Time rounded to the nearest 15 minutes, and one flow value. The time-of-day formula has not beaten the historical center near midnight, so the guess uses that honest baseline.
Historical testing improved median-flow error over ordinary climatology by approximately 41% at 7 days, 28% at 30 days and 11% at 90 days. Winter snowpack error averaged about 3.2 inches of water at 30 days and 6.7 inches at 90 days. Accuracy and uncertainty vary sharply by season and forecast horizon.
The flow ranges are now calibrated separately by horizon: historical coverage was about 81% at 7 days and about 80% at 14, 30, 60 and 90 days. A later range can occasionally be narrower because it lands in a historically steadier season; the range follows observed error rather than being forced to widen.
Ask the Black Hat uses 12 independently trained forecast horizons from one through 90 days and blends between the nearest pair for the date a visitor selects. Chronological backtesting covered 77,074 future dates; its experimental ranges were calibrated to contain the result about four times out of five within each horizon band. The displayed odds are rounded to the nearest five percent because a 73.6% crystal ball would be taking itself much too seriously.
The eight fixed predictions do not share one magic formula. Archived NOAA one-month and three-month outlooks are used on the first-400, May–June strength and spring-peak cards where they improved the relevant historical tests. The rebuilt Peak Snowpack card uses the February–April outlook only as a warm-and-dry guardrail when local snow is already unusually low. The top flow estimates, hydrograph, threshold table, opening-day and tubing-day cards kept their stronger no-NOAA versions.
NOAA's archived 8–14-day outlook tied the existing eight-card method rather than beating it. In July–September, adding the one- or three-month climate outlook slightly worsened flow error and tubing-range probabilities, so neither was added to the Monsoon Flow Watch. The Watch still uses the National Weather Service's live seven-day headwater precipitation forecast as short-range storm context.
For the eight-card lineup, rolling tests on water years 2014–2025 produced 73 wins in 96 development card-seasons. A method selection frozen after 2020 then won 30 of 40 cards in 2021–2025—the same total as the no-NOAA lineup, although NOAA reduced several important numeric errors. The new Peak Snowpack method retained its eight wins in 12 tests while reducing typical error from about 5.6 to 4.6 inches; its locked 2026 miss remains unchanged on the scoreboard. These are experiments, not promises.
Ask the Black Hat!
Pick a date during the next 90 days and the contraption will estimate the odds that San Juan River flow lands in Pagosa’s general tubing and rafting ranges.
Pick a future date to wake it up.
Once the Lab finishes loading, your answer will use the newest river and snowpack inputs.
These are odds of reaching flow thresholds—not a promise that a trip will operate. Water temperature, debris, access, weather and the River Crew’s daily call still decide the lineup.
U.S. Geological Survey Pagosa gauge supplies river flow. Upper San Juan and Wolf Creek Summit Snow Telemetry stations are the two primary snow sites inside the contributing headwaters. Middle Creek and Lily Pond are outside the drainage and are used only as regional storm-pattern checks on the exact cards where chronological testing improved. NOAA Climate Prediction Center one- and three-month outlooks act as low-weight climate context or guardrails on four seasonal cards; they are excluded everywhere they failed to improve testing. National Weather Service seven-day forecast data feeds the live Watch. Colorado Basin River Forecast Center guidance remains the official forecast benchmark. Josh Kurz’s South San Juans streamflow writing provides an independent human-powered perspective. Beartown, Grayback and the combined outside-station set were tested and rejected.
Enough Guessing?
Trade the sparks and speculation for actual San Juan River flow, water temperature, short-term forecasts and today’s river options.
FAQs
Stuff you want to know about Pagosa Outside's Black Hat Hydrology Lab Experiment
The Black Hat Hydrology Lab is Pagosa Outside’s public-facing river-forecasting experiment. It combines historical San Juan River flows, headwater snowpack, current conditions and weather information to make longer-range flow estimates and seasonal predictions. It is part river science, part local experience and part ongoing attempt to make better educated guesses than we used to make over coffee.
The name is a wink, not a confession. In our version, “black hat” means taking public river, snowpack, weather and climate information traditionally scattered across separate sources, hacking it together in unconventional ways and testing whether it can produce a better-informed home-field advantage.
We are not breaking into government computers or hiding the methods: every source is public, every official guess is locked and every miss stays on the scoreboard. The only thing we are trying to outsmart is Mother Nature—and she still has the administrator password.
The River Desk shows current measurements, published government forecasts, snowpack, weather and today’s river options. The Black Hat Lab goes beyond those published forecasts and experiments with what might happen 30 to 90 days from now or later in the river season.
The Lab’s predictions are independent Pagosa Outside experiments—not official government forecasts. They should be treated as entertainment rather than instructions for river use, travel or business decisions. Use the River Desk for current conditions and published short-term forecasts.
The Lab estimates San Juan River flow 30, 60 and 90 days ahead and displays an experimental 90-day hydrograph. It calculates the probability of reaching Pagosa’s general tubing and rafting-flow ranges, operates a Live Seasonal Watch and publishes eight fixed predictions about important events during the water year.
Those seasonal predictions cover headwater snowpack, the beginning and strength of spring runoff, peak flow, rafting-flow days, Tubing Season Opening Day and tubing-range days.
The honest answer is: promising, but still experimental. In historical testing, the seasonal prediction methods scored 73 hits in 96 development tests. A later five-season test—using methods selected without seeing those results—scored 30 hits in 40 predictions. That works out to roughly three correct calls out of four.
The 30-day flow model reduced its typical error by approximately 28% compared with using historical averages alone, while the 90-day model improved by approximately 11%. The experimental ranges contained the actual result about four times out of five during historical testing.
Accuracy still changes considerably by season, forecast distance and type of prediction. One unusual storm, heat wave or stubborn snowpack can embarrass the entire contraption. That is why the scoreboard displays every public hit, miss and pending result instead of showing only the guesses that make us look clever.
The Lab uses river flow from the U.S. Geological Survey gauge at Pagosa Springs; snowpack from the Upper San Juan and Wolf Creek Summit Snow Telemetry stations; temperature and precipitation forecasts from the National Weather Service; selected climate outlooks from the National Oceanic and Atmospheric Administration; and historical guidance from the Colorado Basin River Forecast Center.
Middle Creek and Lily Pond may be used as regional storm-pattern checks on predictions where historical testing showed an improvement. Other stations—including Beartown and Grayback—were tested but did not improve the Pagosa predictions enough to earn a regular seat at the workbench.
The Lab does not ask a chatbot to stare into the river and invent a number. Its predictions come from repeatable formulas, historical comparisons and backtested models. The calculations run automatically, but the targets, inputs, scoring rules and operating thresholds were developed around the actual questions Pagosa river managers and guides try to answer.
The Lab compares current river flow, recent movement, seasonal timing, headwater snowpack, temperature and precipitation with similar conditions from previous water years. It produces a center estimate and an experimental range for each future date. The historical average is displayed as a baseline, not as a second prediction.
The center estimate is the Lab’s single best guess. The experimental range shows a broader group of reasonably plausible outcomes and is calibrated to contain the actual result about four times out of five during historical testing. It is not a best-case and worst-case limit, and the river is under no contractual obligation to remain inside it.
Each horizon is tested separately. A 60-day range can occasionally be wider than a 90-day range when the 60-day date lands during a volatile runoff or seasonal transition and the 90-day date lands during historically steadier flow.
The probability table estimates the chance that flow on each forecast date will fall within Pagosa’s general operating ranges:
30–399 cubic feet per second: tubing range
300 cubic feet per second or higher: possible Splash & Dash rafting flow
600 cubic feet per second or higher: possible Mesa Canyon rafting flow
800 cubic feet per second or higher: possible West Fork rafting flow
These are flow probabilities—not promises that a particular trip will operate. Water temperature, weather, debris, access and current river conditions still influence Pagosa Outside’s daily lineup.
The Watch changes jobs as the water year progresses. During winter, Spring Runoff Watch follows headwater snowpack development and early runoff potential. As spring runoff begins rising, Predict the Peak Runoff Watch continuously updates the Lab's best guess for the peak date, Mountain Time and flow. After a sustained decline indicates that the spring crest has likely passed, it becomes Tubing Opening Day Watch and looks for the first complete 24 hours with every available Pagosa gauge reading from 30 through 399 cubic feet per second. After the tubing transition, Monsoon Flow Watch follows possible rain-driven tubing boosts, low-water recoveries and returns to rafting flow. After Labor Day, it becomes Bonus Flow Watch.
Unlike the eight official predictions, the Live Seasonal Watch updates whenever the Lab runs and new information becomes available. Its changing estimates never rewrite the fixed guesses on the scoreboard.
The eight official predictions are:
First persistent headwater snowpack
Peak headwater snowpack
First 400-plus cubic-feet-per-second day
May–June river strength
Spring peak date, time and flow
Rafting-flow days from May 1 through Labor Day
Tubing Season Opening Day
Tubing-flow days from May 1 through Labor Day
Each question uses its own combination of snowpack, river flow, weather, climate outlooks, seasonal timing and historical comparisons. Predicting peak snowpack is not the same problem as predicting spring peak flow or tubing days, so one magic formula would not be very magical.
Each official call also has a fixed issue date selected to balance useful lead time with historical accuracy. Once published, the guess is locked and cannot be quietly changed when better information arrives.
Every prediction has a measurable finish line and scoring rule established before the outcome is known. A hit earns one point for the Black Hat Lab. A miss earns one point for Mother Nature. Unfinished predictions remain pending.
The scoreboard follows the October 1 through September 30 water year and divides the season into four quarters. A new game begins each October, while the previous season’s record remains visible. The Lab does not get to quietly drag its misses behind the shed.
A retrospective season recreates earlier predictions using only information that would have been available on each stated issue date. The model is trained on previous years and is not allowed to peek at what happened later. Retrospective results are clearly labeled because they were historical tests—not predictions publicly posted at the time.
Sometimes. Archived National Oceanic and Atmospheric Administration one-month and three-month outlooks improved selected predictions, including the first 400-plus cubic-feet-per-second day, May–June river strength and spring peak flow.
Those outlooks did not improve every part of the Lab. They are excluded from the 30-, 60- and 90-day flow estimates, hydrograph, operating-threshold probabilities and Monsoon Flow Watch because historical testing found that the other methods performed as well or better. The Live Seasonal Watch still uses the National Weather Service’s seven-day headwater forecast for short-range weather and storm information.
More data does not automatically produce a better prediction, so every new ingredient has to earn its way into the formula.
Absolutely. This is a continuing public experiment, and the build number shows which version is running. New water years create new tests, misses expose weak formulas and better data sources may improve individual predictions.
A method changes only when backtesting shows a real improvement without peeking at the answer. Past guesses and results remain visible so a new formula cannot rewrite history or polish an old miss into a hit.
Because once we put the whole river story onto one desktop, somebody asked the dangerous question: What happens after the official forecast ends—and can we guess it better than usual?
The Lab combines 25 years of river and snowpack history with current conditions, weather forecasts, tested formulas and the PO River Crew’s accumulated hunches to make longer-range and seasonal predictions. It is where actual science, river-guide instinct and old-timer theories are forced to share one questionable workbench while Mother Nature controls the scoreboard.
We built it to find out whether our educated guesses can beat ordinary historical averages—and because arguing about peak runoff at the coffee shop was not producing enough charts.