· metodología · probabilidades · MECE · análisis
Cómo se le pone una probabilidad a un caso UAPHow to put a probability on a UAP case
La conversación pública sobre fenómenos anómalos casi siempre colapsa en dos casillas: o «son extraterrestres» o «no hay nada». Esa binariedad es cómoda para discutir y pésima para analizar, porque obliga a tratar igual a un planeta confundido con una nave y a un incidente con triple registro de sensores. UAP Codex usa una herramienta distinta, prestada del análisis de inteligencia: en lugar de un veredicto, a cada caso se le asigna una distribución de probabilidad. No «qué fue», sino «cómo se reparte la creencia razonable sobre qué fue».
La regla es que cada caso de incidente reparte exactamente el 100% de su probabilidad entre seis narrativas que son, a la vez, mutuamente excluyentes y exhaustivas —lo que en análisis se llama una partición MECE—. Las seis: mundano/natural (identificación errónea, fenómenos atmosféricos), humana clasificada (un programa propio o aliado no reconocido), adversaria (tecnología de otro Estado), no-humano encubierto (algo no-humano que un Estado conoce u oculta), no-humano abierto (algo no-humano que nadie controla, al estilo de las hipótesis de Vallée) e indeterminado. Que sumen 100% y no se solapen es lo que vuelve honesto el ejercicio: no puedes inflar una hipótesis sin restarle a otra.
El paso clave es la agregación. Si cada caso es una distribución, ¿cómo se habla del corpus entero? Sumando: el número esperado de casos que caen en cada narrativa es la suma, sobre todos los casos, de la probabilidad que cada uno le asigna a esa narrativa. Como es una esperanza matemática —una operación lineal— el resultado es válido y comparable entre narrativas aunque los casos estén correlacionados entre sí. El total se reparte en 200 piezas y cada narrativa se lleva su fracción. Eso permite responder preguntas que el debate binario no puede: ¿pesa más lo mundano o lo indeterminado? ¿qué tan grande es realmente la franja no-humana cuando se mira todo junto, y no solo el caso más famoso?
La forma de los datos, a junio de 2026 sobre el corpus completo, desinfla por igual a entusiastas y a escépticos duros. El peso agregado se inclina con claridad hacia lo prosaico: cerca de la mitad del total cae en explicaciones mundanas o naturales. Alrededor de un cuarto queda, con honestidad, como indeterminado —ni resuelto ni forzado a una conclusión—. Y la fracción no-humana, lejos de dominar el debate como sugiere la cultura popular, ronda apenas una sexta parte del total. El desglose vivo y exacto, que cambia a medida que el corpus crece y se reclasifica, se publica de forma transparente en la página de probabilidades del sitio.
Conviene subrayar qué no es esto. No son frecuencias calibradas ni mediciones: son juicios analíticos estructurados, en la tradición del estándar ICD-203 de la comunidad de inteligencia, que separa la probabilidad de un evento de la confianza en la evidencia que la sustenta. Comparabilidad no es verdad: que dos casos sean comparables en la misma escala no garantiza que la escala acierte. Los casos-documento (memos, informes, leyes) se excluyen del reparto, porque la pregunta «qué era el objeto» no se aplica a un papel. Y cada número es revisable: si aparece una fuente primaria nueva, el reparto de ese caso cambia, y con él el agregado.
El valor de tratar el fenómeno así no es llegar a una respuesta final, sino imponerse una disciplina. Obliga a explicitar la incertidumbre en vez de esconderla detrás de un titular, a no robarle peso a una hipótesis para regalárselo a otra, y a que cualquiera pueda auditar el razonamiento caso por caso. En un campo donde casi todo el mundo ya decidió de antemano lo que va a creer, poner una probabilidad —y dejar que los 200 casos hablen en conjunto— es una forma modesta pero exigente de honestidad.
Public conversation about anomalous phenomena almost always collapses into two boxes: either «they're extraterrestrials» or «there's nothing here.» That binary is comfortable for arguing and terrible for analysis, because it forces you to treat a planet mistaken for a craft and an incident with triple sensor recording exactly the same way. UAP Codex uses a different tool, borrowed from intelligence analysis: instead of a verdict, each case is assigned a probability distribution. Not «what it was,» but «how reasonable belief about what it was should be divided.»
The rule is that each incident case splits exactly 100% of its probability across six narratives that are, at once, mutually exclusive and exhaustive —what analysts call a MECE partition. The six: mundane/natural (misidentification, atmospheric phenomena), classified human (an unacknowledged domestic or allied program), adversarial (another state's technology), covert non-human (something non-human that a state knows of or hides), open non-human (something non-human that no one controls, in the vein of Vallée's hypotheses), and indeterminate. That they sum to 100% and do not overlap is what keeps the exercise honest: you cannot inflate one hypothesis without subtracting from another.
The key step is aggregation. If each case is a distribution, how do you talk about the whole corpus? By summing: the expected number of cases falling under each narrative is the sum, over all cases, of the probability each one assigns to that narrative. Because it is a mathematical expectation —a linear operation— the result is valid and comparable across narratives even if the cases are correlated. The total is divided into 200 pieces and each narrative takes its fraction. That answers questions the binary debate cannot: does the mundane outweigh the indeterminate? How large is the non-human band really, when you look at everything together rather than just the most famous case?
The shape of the data, as of June 2026 over the full corpus, deflates enthusiasts and hard skeptics alike. The aggregate weight leans clearly toward the prosaic: close to half the total falls under mundane or natural explanations. About a quarter remains, honestly, indeterminate —neither resolved nor forced into a conclusion. And the non-human fraction, far from dominating the debate as popular culture suggests, is barely a sixth of the total. The live, exact breakdown, which shifts as the corpus grows and is reclassified, is published transparently on the site's probabilities page.
It is worth stressing what this is not. These are not calibrated frequencies or measurements: they are structured analytical judgments, in the tradition of the intelligence community's ICD-203 standard, which separates the probability of an event from confidence in the evidence behind it. Comparability is not truth: that two cases are comparable on the same scale does not guarantee the scale is right. Document cases (memos, reports, laws) are excluded from the split, because the question «what was the object» does not apply to a piece of paper. And every number is revisable: if a new primary source appears, that case's split changes, and with it the aggregate.
The value of treating the phenomenon this way is not reaching a final answer, but imposing a discipline. It forces you to make uncertainty explicit instead of hiding it behind a headline, not to steal weight from one hypothesis to hand it to another, and to let anyone audit the reasoning case by case. In a field where almost everyone has already decided in advance what they will believe, putting a probability on it —and letting all 200 cases speak together— is a modest but demanding form of honesty.