#EconML
The LOST-STATS page on causal forest has a Python implementation that apparently no longer works. Anyone familiar with econml and able to help update the code? I can handle the page update if you can fix the code lost-stats.github.io/Machine_Lear...
Causal Forest
lost-stats.github.io
February 18, 2025 at 9:14 PM
And if you're looking for something with Python code:

* amzn.to/4fS1bfN (Causal Inference & Discovery by yours truly)
* amzn.to/41993VH (Causal Inference by Matheus Facure)

#CausalSky
December 1, 2024 at 5:41 PM
As a default, it seems reasonable to go w/ DoubleML if you're sure that DML is a methodology you want to use & there's a specific estimator you want to use impl. in the pkg

EconML seems a better choice when you're experimenting with various estimators &/or want access to the DoWhy/EconML feats

3/3
November 18, 2024 at 11:32 PM
Anyone know of anything like "A list of Python packages an applied economist should install to allow Python to do what stata does" - just a nice list of all the great packages people have made to get estimation tools in Python. I'm thinking e.g. pyfixest, binsreg, econml, etc. Any leads?
November 8, 2024 at 5:59 PM
My PR to the #EconML #PyWhy #opensource #causalai project was merged! 🎉 I made a small contribution by allowing a flexible choice of evaluation metric for scoring both the first stage and final stage models in Double Machine Learning (#DML).
#CausalInference #machinelearning #datascience
July 10, 2025 at 2:55 PM
Hi Jake,

Yes, I am. EconML is a part of a broader ecosystem and it allows you to easily compare different types of estimators (not only DML). DoubleML focuses on DML specifically.

EconML gives you access to all the functionalities of the DoWhy/EconML ecosystem, including refutation tests,...

1/n
November 18, 2024 at 11:23 PM
you don't know how strongly i agree with you; spending weeks F12-ing econml classes to figure out where the linalg is done is excellent anti-OOP indoctrination. DRY principle is good, but within reason.
November 15, 2024 at 10:44 PM
Introductory post! We are causalwizard.app , a website that helps researchers learn about and adopt Causal Inference. The app has a database of about 100 practical causality articles and includes tools from DoWhy, EconML and StatsModels.
#CausalSky #EpiSky #MachineLearning #DataScience #Statistics
Causal Wizard
Explore cause and effect in historical data; predict the effects of counterfactual scenarios and other interventions using the latest Causal Inference methods and machine learning tools, in an onlin...
causalwizard.app
November 13, 2024 at 12:13 AM
Made my first PR on an #OpenSource project. Before now, all the projects I worked with had everything I needed. The PR is just a small suggestion to improve memory performance in #EconML (#causalml) - we'll see what the repo maintainers think of my hackery, haha...
January 22, 2025 at 3:55 PM
...evaluation module and more.

DoubleML has more use case specific implementations of the DML framework, a great sensitivity analysis module (a feat. also available in EconML, but less advanced (at least this is how it was some time ago)), and more.

2/n
November 18, 2024 at 11:28 PM
oh one other change: econml now (rightfully) distinguishes between covariates used to fit E[Y|X], E[T|X] and the ones used to estimate CATE functions. so i got it to run by passing the same matrix twice (i.e. hetfx by all covars - grf default) but that is probably not what anyone should ever do
February 18, 2025 at 10:19 PM
The new version of EconML offers some cool new features (stay tuned, we'll cover them soon) and a couple of very interesting code bases came to life with recent NeurIPS papers (we'll discuss some of them this year as well).

3/n
January 1, 2024 at 9:56 AM
arXiv📈🤖
A Large-Scale Empirical Comparison of Meta-Learners and Causal Forests for Heterogeneous Treatment Effect Estimation in Marketing Uplift Modeling
By Singh
April 8, 2026 at 4:58 AM
But do we have to intentionally corrupt an AI assistant to get misleading results?

Earlier this year, a Stanford team led by Vasilis Syrgkanis (among other contexts, you might know him as one of the core EconML developers) tested ChatGPT 5.3 on a variety of causal problems.

5/
September 16, 2026 at 9:30 AM
Machine Learning in Clojure with libpython‑clj: Unlocking Causal Insights Using Microsoft’s EconML [Series 3] “Beyond A/B testing: causal inference meets functional programming.” In the fir...

Origin | Interest | Match
December 23, 2025 at 4:41 PM
If you want to take your first (or second) steps in causality, without feeling overwhelmed, here are three things you can do:

1. I wrote an accessible book on this topic, you can get a copy here: amzn.to/49b9b7Y

8/n 👇🏼
Causal Inference and Discovery in Python: Unlock the secrets of modern causal machine learning with ...
Causal Inference and Discovery in Python: Unlock the secrets of modern causal machine learning with DoWhy, EconML, PyTorch and more [Molak, Aleksander, Jaokar, Ajit] on Amazon.com. *FREE* shipping on ...
amzn.to
February 2, 2024 at 9:38 AM

To learn more about the assumptions needed for causal discovery see Chapter 13 of my recent book (amzn.to/3upL3zh ) where I summarize them for you.

7/7
Causal Inference and Discovery... by Molak, Aleksander
Causal Inference and Discovery in Python: Unlock the secrets of modern causal machine learning with DoWhy, EconML, PyTorch and more [Molak, Aleksander, Jaokar, Ajit] on Amazon.com. *FREE* shipping on ...
amzn.to
November 20, 2023 at 11:53 AM
The Three Revolutions Nobody Connected

The AI industry is living through three simultaneous revolutions -- in separate silos, blissfully unaware of each other.

Revolution 1: Causal AI. Judea Pearl gave us do-calculus and structural causal models. Companies like CausaLens raised $45M+ to bring […]
Functors, do-calculus, and Phase Transitions: The Unification Layer AI Companies Don't See Coming
## The Three Revolutions Nobody Connected The AI industry is living through three simultaneous revolutions -- in separate silos, blissfully unaware of each other. **Revolution 1: Causal AI.** Judea Pearl gave us do-calculus and structural causal models. Companies like CausaLens raised $45M+ to bring causality to enterprise. Microsoft shipped DoWhy and EconML. The promise: AI that answers "why," not just "what." **Revolution 2: Categorical AI.** Symbolica AI and Google DeepMind co-published the landmark ICML 2024 paper arguing that category theory is "an algebraic theory of all architectures." Quantinuum's Bob Coecke group formalized causal models as string diagrams. The promise: AI with compositional guarantees -- if the parts are correct, the whole is provably correct. **Revolution 3: Complexity-Aware AI.** Jiang Zhang's group at Beijing Normal University built the Neural Information Squeezer and SVD-based Causal Emergence theory. Harvard/NTT showed transformers exhibit phase transitions following percolation theory. The promise: AI that understands emergence -- when the whole becomes more than the sum of its parts. Here's the problem: **no company, no lab, no framework has unified all three.** CausaLens does causality without composition. Symbolica does composition without emergence. Beijing Normal does emergence without structural causality. They're all building one leg of a three-legged stool. This article argues that the convergence of these three fields -- what I call **Compositional Causal Emergence (CCE)** -- is not an academic curiosity. It's AI's next paradigm shift, and the window for first-mover advantage is open right now. --- ## The Blind Spots: What Each Silo Can't See ### Causal AI's Frozen Map Problem Every causal AI system today assumes a **fixed causal graph**. You draw a DAG, label the nodes, define the arrows, then run do-calculus on it. This works beautifully in controlled environments. But real-world systems don't have fixed causal structures. In financial markets, the causal relationships between variables shift during crises. In healthcare, a patient's causal response to treatment changes as comorbidities emerge. In climate systems, feedback loops create entirely new causal pathways that didn't exist a decade ago. Current Causal AI is a GPS navigation system that draws the map perfectly but can't handle the fact that the roads keep changing. ### Category Theory's Earthquake Problem Categorical approaches deliver stunning compositional guarantees. If module A is verified and module B is verified, their composition A + B is provably correct. This is the "correct-by-construction" dream. But complex systems violate this assumption constantly. The whole is more than the sum of its parts. Non-linear interactions, feedback loops, and emergent behaviors mean that perfectly verified components can produce unpredictable system-level behavior. A building made of flawless LEGO bricks still needs earthquake engineering. ### Complexity Science's Forensics Problem Causal emergence research can detect that macro-level dynamics are causally stronger than micro-level dynamics. Zhang's SVD framework (2024, npj Complexity) made this computationally tractable. But complexity science can find the earthquake -- it just can't tell you which column broke, or what would have happened if you'd reinforced it. Without structural causal models, emergence detection is observation without intervention. Without compositional guarantees, emergence analysis doesn't scale to modular systems. --- ## The Convergence: Three Operations, One Chain CCE combines three fundamental operations into a single executable chain: **COMPOSE** (Category Theory) -- Combine causal sub-models with structural guarantees using Markov categories and monoidal composition. **INTERVENE** (Causal AI) -- Calculate the effect of interventions in the composed model using SCMs and do-calculus, including counterfactual reasoning. **EMERGE** (Complexity Science) -- Measure whether macro-level causal effects are stronger than micro-level effects using effective information and SVD decomposition. The critical innovation: these three operations can be **composed on top of each other** : > EMERGE after INTERVENE after COMPOSE Meaning: compose the parts, intervene in the whole, then measure whether the intervention changed the emergent structure itself. This chain has never been formalized in any existing framework. --- ## The Evidence: This Convergence Is Already Beginning The building blocks exist. They just haven't been assembled: **Lorenz & Tull (2026, Quantinuum)** formalized causal abstraction -- moving between micro and macro causal descriptions -- as natural transformations in category theory. This directly bridges causality and composition. Missing piece: no complexity metrics, no emergence detection. **Krasnovsky (2025)** published the closest existing work: sheaf-theoretic causal emergence combining categorical structures (sheaves), causal metrics (effective information), and complexity (emergence) for distributed systems analysis. Missing piece: applied to microservices and power grids, not yet to AI/ML models. **Zhang et al. (2024, npj Complexity)** eliminated the coarse-graining selection problem in causal emergence using SVD of transition matrices. This makes emergence practically computable. Missing piece: no compositional framework, doesn't scale to modular systems. **Gavranovic et al. (2024, ICML, Symbolica + DeepMind)** showed all neural network architectures can be expressed in the universal algebra of monads. Missing piece: no causality, no complexity -- purely structural. **Xiao et al. (2026)** proved that categorical causal formalization improves LLM performance by 4.9% over state-of-the-art. This isn't theory anymore -- it's measurable performance gain. **Nye (2025)** established a bijection between Lawvere theories and neural architectures, producing networks that are **mathematically incapable** of violating specified logical constraints. Correct-by-construction AI is not a dream; it's demonstrated. --- ## Why This Matters Now: The Regulatory Forcing Function The EU has inadvertently created the most powerful market pull for this convergence. **AI Act Article 73** requires high-risk AI providers to demonstrate a "causal link" between their system and serious incidents. Not correlation. Not feature importance scores. A causal link. **The AI Liability Directive** introduces a "rebuttable presumption of causality." If an AI system fails to comply with AI Act rules and damage occurs, courts will **presume the AI caused the damage** until the provider proves otherwise. This makes Causal AI not a luxury but a **defensive shield**. And here's what no compliance vendor has realized: SHAP and LIME -- the tools everyone uses for "explainability" -- measure correlation, not causation. They show that a loan rejection correlates with a ZIP code, but they cannot prove whether that ZIP code is acting as a proxy for race or income. Only structural causal models can make that distinction. The company that delivers compositional causal explanations -- explanations that are mathematically guaranteed to be consistent across system components, that can prove causal links, and that can detect emergent risks before they materialize -- doesn't just solve a compliance problem. It sets the **de facto technical standard** for AI Act compliance. --- ## The White Space: What Nobody Has Built The market map tells a clear story: Capability | Who's Doing It | What's Missing ---|---|--- Causal AI | CausaLens, Microsoft DoWhy | No composition, no emergence Categorical AI | Symbolica, Quantinuum | No causality metrics, no emergence Complexity/Emergence | Beijing Normal, Santa Fe Institute | No structural causality, no composition **All Three Combined** | **Nobody** | **The entire framework** No product, no tool, no startup occupies the center of this triangle. The entity that builds a **"Compositional Causal Emergence Engine"** -- a system that composes verified causal modules, runs interventions across scales, and monitors emergent behavior in real time -- occupies an entirely new category. --- ## Five Multiplier Effects **1. Compliance Arbitrage.** First-mover in compositional causal explanations becomes the technical standard for EU AI Act compliance. Estimated compliance cost reduction: 60-80% vs. ad hoc approaches. **2. AI Safety Beyond Empiricism.** Current AI safety is either empirical (red-teaming) or statistical (adversarial testing). CCE adds structural guarantees: mathematical proofs that certain behaviors cannot emerge under specified conditions. This is a qualitatively different safety layer. **3. Interpretability as Legal Evidence.** String diagrams aren't just pretty math. They produce visual causal audit trails that can serve as legal evidence in liability proceedings -- a "causal chain certificate" for the courtroom. **4. Cross-Domain Transfer.** Functorial mappings preserve causal structure across domains. A causal model of drug interactions can be formally transferred to model financial contagion -- if the structures are isomorphic, the transfer is mathematically guaranteed. **5. Quantum-Ready Architecture.** Markov categories encompass both classical and quantum systems. A CCE framework transitions to quantum causal inference with zero architectural changes when quantum hardware matures. --- ## The Call to Action Newton had three laws. Einstein unified them into spacetime. AI has three revolutions -- causality, compositionality, complexity -- waiting for their unification moment. The pieces are on the table. Lorenz and Tull gave us categorical causal abstraction. Zhang gave us computable emergence. Gavranovic gave us categorical deep learning. Krasnovsky showed the triple convergence is possible. What's missing is the engineering integration and the will to build it. The white space is open. The regulatory wind is at your back. And the first team that writes AI's equivalent of Maxwell's equations -- unifying seemingly separate research programs into a single framework -- won't just publish a paper. They'll define the next decade of AI. ---
www.mesutaydin.link
March 14, 2026 at 11:28 AM
My journey was something like this

1. Jonas Peters introductory lectures
2. Judeal Pearl's Book of Why, causality and Primer
3. Rubin's PO, Angrist, Imbens work on IV
4. Hernan and Robin's What If book
5. Coding causal models with R/Python (dagitty, doWhy, econML)
6. Causality in LLMs

#CausalSky
January 10, 2025 at 4:52 AM