kth.sePublications KTH
Change search
Link to record
Permanent link

Direct link
Publications (9 of 9) Show all publications
Agram, N., Benth, F. E., Pucci, G. & Rems, J. (2026). A deep learning approach to renewable capacity installation under jump uncertainty. Mathematics and Computers in Simulation, 250, 152-177
Open this publication in new window or tab >>A deep learning approach to renewable capacity installation under jump uncertainty
2026 (English)In: Mathematics and Computers in Simulation, ISSN 0378-4754, E-ISSN 1872-7166, Vol. 250, p. 152-177Article in journal (Refereed) Published
Abstract [en]

We study a stochastic model for the installation of renewable energy capacity under demand uncertainty and jump driven dynamics. The system is governed by a multidimensional Ornstein–Uhlenbeck (OU) process driven by a subordinator, capturing abrupt variations in renewable generation and electricity load. Installation decisions are modeled through control actions that increase capacity in response to environmental and economic conditions. We consider two distinct solution approaches. First, we implement a structured threshold based control rule, where capacity is increased proportionally when the stochastic capacity factor falls below a fixed level. This formulation leads to a nonlinear partial integro-differential equation (PIDE), which we solve by reformulating it as a backward stochastic differential equation with jumps. We extend the DBDP solver in Huré et al. (2020) to the pure jump setting, employing a dual neural network architecture to approximate both the value function and the jump sensitivity. Second, to benchmark the performance of the first approach, we propose a deep control algorithm that directly learns a state and time dependent feedback threshold policy by minimizing the expected cost functional using neural networks. This approach preserves the threshold based structure while allowing the threshold to adapt dynamically to the evolving system state, enabling more flexible and responsive interventions. Numerical experiments on calibrated data highlight the strengths of both methods. While the threshold based BSDE approach offers interpretability and tractability, the deep control strategy achieves improved performance through flexibility in capacity allocation. Together, these tools provide a robust framework for decision support in long term renewable energy expansion under uncertainty.

Place, publisher, year, edition, pages
Elsevier BV, 2026
Keywords
BSDEs, Capacity installation, Deep learning, Jump diffusion, Optimal stochastic control, Renewable energy expansion, Threshold control
National Category
Computational Mathematics Control Engineering
Identifiers
urn:nbn:se:kth:diva-384812 (URN)10.1016/j.matcom.2026.06.024 (DOI)2-s2.0-105042678051 (Scopus ID)
Note

Not duplicate with diva 2025302

QC 20260703

Available from: 2026-07-03 Created: 2026-07-03 Last updated: 2026-07-03Bibliographically approved
Agram, N. & Pucci, G. (2026). Deep BSVIEs parametrization and learning-based applications. Neural Networks, 198, 108712
Open this publication in new window or tab >>Deep BSVIEs parametrization and learning-based applications
2026 (English)In: Neural Networks, ISSN 0893-6080, E-ISSN 1879-2782, Vol. 198, p. 108712-Article in journal (Refereed) Published
Abstract [en]

We study the numerical approximation of backward stochastic Volterra integral equations (BSVIEs) and their reflected extensions, which naturally arise in problems with time inconsistency, path dependent preferences, and recursive utilities with memory. These equations generalize classical BSDEs by involving two dimensional time structures and more intricate dependencies. We begin by developing a well posedness and measurability framework for BSVIEs in product probability spaces. Our approach relies on a representation of the solution as a parametrized family of backward stochastic equations indexed by the initial time, and draws on results of Stricker and Yor to ensure that the two parameter solution is well defined in a joint measurable sense. We then introduce a discrete time learning scheme based on a recursive backward representation of the BSVIE, combining the discretization of Hamaguchi and Taguchi with deep neural networks. A detailed convergence analysis is provided, generalizing the framework of deep BSDE solvers to the two dimensional BSVIE setting. Finally, we extend the solver to reflected BSVIEs, motivated by applications in delayed recursive utility with lower constraints.

Place, publisher, year, edition, pages
Elsevier BV, 2026
Keywords
BSVIEs, Deep learning, Neural network solvers, RBSVIEs, Stricker-Yor measurability
National Category
Computational Mathematics Probability Theory and Statistics Other Mathematics
Identifiers
urn:nbn:se:kth:diva-378534 (URN)10.1016/j.neunet.2026.108712 (DOI)001695064200001 ()41707458 (PubMedID)2-s2.0-105032249435 (Scopus ID)
Note

Not duplicate with DiVA 1982180

QC 20260327

Available from: 2026-03-27 Created: 2026-03-27 Last updated: 2026-03-27Bibliographically approved
Pucci, G. (2026). Deep Learning and Optimal Stochastic Control with Applications. (Doctoral dissertation). Stockholm: KTH Royal Institute of Technology
Open this publication in new window or tab >>Deep Learning and Optimal Stochastic Control with Applications
2026 (English)Doctoral thesis, comprehensive summary (Other academic)
Abstract [en]

This thesis brings together theoretical advances in stochastic optimal control and modern deep learning techniques, with particular emphasis on applications in environmental and energy systems. The first group of contributions investigates optimal control from a theoretical perspective, developing new results and illustrating their relevance through real world applications. The second part explores deep learning methods for solving stochastic differential equations and control problems that are analytically intractable.

We begin by studying impulse control problems for conditional McKean--Vlasov jump diffusions, extending the classical verification theorem to the setting in which the state dynamics depend on their conditional distribution. We then examine an optimal control problem for pollution growth on a spatial network, formulated in a deterministic framework but capturing how environmental policies propagate across interconnected geographical regions. Finally, we develop a model for investment in renewable energy capacity under uncertainty, characterising how optimal installation strategies change in response to fluctuations in energy demand and production. These contributions show how stochastic control can be used to address pressing challenges in environmental regulation and energy planning.

The second line of research focuses on deep learning methods for backward stochastic differential equations (BSDEs) and related formulations, together with direct machine learning approaches for high-dimensional stochastic control. Specifically, we solve Dynkin games by reformulating them as doubly reflected BSDEs, enabling the computation of optimal stopping strategies in energy market contracts. We further develop a deep learning solver for backward stochastic Volterra integral equations (BSVIEs), extending neural BSDE methods to systems with memory. In addition, we propose a machine learning framework for renewable capacity investment under jump uncertainty, treating the problem both through a direct control learning strategy and through a newly developed solver for pure jump BSDEs.

Overall, this thesis lies at the intersection of rigorous mathematical analysis and machine learning-based approaches to stochastic optimal control. On the one hand, we show how careful modeling and theoretical results enable the formulation and study of complex, realistic control problems; on the other hand, we demonstrate how modern machine learning techniques provide powerful tools for solving these problems efficiently. The applications are motivated by urgent questions in environmental and energy sustainability.

Abstract [sv]

Denna avhandling förenar teoretiska framsteg inom stokastisk optimal styrning med moderna djupinlärningsmetoder, med särskild tonvikt på tillämpningar inom miljö- och energisystem. Den första gruppen av bidrag undersöker optimal styrning ur ett teoretiskt perspektiv, utvecklar nya resultat och visar dess relevans genom praktiskt motiverade exempel. Den andra delen behandlar djupinlärningsmetoder för att lösa stokastiska differentialekvationer och styrproblem som annars är analytiskt oöverskådliga.

Vi börjar med att studera impulskontrollproblem för betingade McKean–Vlasov-hoppdiffusioner och utvidgar den klassiska verifikationssatsen till situationer där systemets dynamik beror på dess betingade fördelning. Därefter analyseras ett optimalt styrproblem för utsläppstillväxt på ett rumsligt nätverk, formulerat deterministiskt men avsett att fånga hur miljöpolitiska beslut sprids över sammanlänkade geografiska regioner. Slutligen utvecklar vi en modell för investering i förnybar energikapacitet under osäkerhet, där vi karakteriserar hur optimala installationsstrategier påverkas av variationer i efterfrågan och produktion. Dessa bidrag visar hur stokastisk styrning kan användas för att hantera centrala frågor inom miljöreglering och energiplanering.

Den andra forskningslinjen fokuserar på djupinlärningsmetoder för bakåtriktade stokastiska differentialekvationer (BSDE:er) och relaterade formuleringar, samt direkta maskininlärningsmetoder för högdimensionella stokastiska styrproblem. Vi löser särskilt Dynkin-spel genom att formulera dem som dubbelt reflekterande BSDE:er, vilket möjliggör beräkning av optimala stoppstrategier i energimarknadskontrakt. Vidare utvecklar vi en djupinlärningsbaserad lösare för bakåtriktade stokastiska Volterra-integralekvationer (BSVIE:er), och utvidgar därmed neurala BSDE-metoder till system med minnesstruktur. Dessutom föreslår vi ett maskininlärningsramverk för investeringar i förnybar kapacitet under hopp-osäkerhet, både genom direkt styrinlärning och genom en ny lösare för rena hopp-BSDE:er.

Sammantaget placerar sig denna avhandling i gränslandet mellan rigorös matematisk analys och maskininlärningsbaserade metoder för stokastisk optimal styrning. Å ena sidan visar vi hur noggrann modellering och teoretiska resultat möjliggör formulering och studie av komplexa, realistiska styrproblem; å andra sidan visar vi hur moderna djupinlärningstekniker ger kraftfulla verktyg för att lösa dessa problem på ett effektivt sätt. Tillämpningarna är motiverade av aktuella och angelägna frågor inom miljömässig och energimässig hållbarhet.

Place, publisher, year, edition, pages
Stockholm: KTH Royal Institute of Technology, 2026. p. 305
Series
TRITA-SCI-FOU ; 2025:71
Keywords
Stochastic optimal control; optimal stopping; Deep learning; Impulse control; McKean--Vlasov dynamics; Jump diffusions; BSDEs; Renewable energy investment; Pollution control; BSVIEs; Machine learning, Stokastisk optimal styrning; Optimalt stopp; Djupinlärning; Impulskontroll; McKean–Vlasov-dynamik; Hoppdiffusioner; BSDE:er; Förnybar energiinvestering; Utsläppskontroll; BSVIE:er; Maskininlärning.
National Category
Computational Mathematics
Research subject
Applied and Computational Mathematics, Mathematical Statistics
Identifiers
urn:nbn:se:kth:diva-375351 (URN)978-91-8106-492-6 (ISBN)
Public defence
2026-02-06, Q2, Malvinas väg 10, Stockholm, 10:00
Opponent
Supervisors
Note

QC 2026-01-13

Available from: 2026-01-13 Created: 2026-01-12 Last updated: 2026-03-20Bibliographically approved
Gozzi, F., Leocata, M. & Pucci, G. (2026). Network-based optimal control of pollution growth. European Journal of Operational Research, 332(3), 1032-1047
Open this publication in new window or tab >>Network-based optimal control of pollution growth
2026 (English)In: European Journal of Operational Research, ISSN 0377-2217, E-ISSN 1872-6860, Vol. 332, no 3, p. 1032-1047Article in journal (Refereed) Published
Abstract [en]

This paper studies a model for the optimal control of pollution diffusion over time and space by a centralized economic agent. The controls are the investments in two types of production: a less polluting (”green”) technology and a more polluting (”brown”) one. The goal is to maximize an intertemporal utility function which takes into account the cost of pollution. The main novelty is the fact that the spatial component has a network structure. Moreover, in such a time-space setting, we analyze the trade-off between the use of green and brown technologies: this is also a novelty in such a setting. Extending methods from previous works, we can explicitly solve the problem in the case of strictly convex or linear pollution costs.

Place, publisher, year, edition, pages
Elsevier BV, 2026
National Category
Economics
Identifiers
urn:nbn:se:kth:diva-382654 (URN)10.1016/j.ejor.2026.03.012 (DOI)001735279300001 ()2-s2.0-105033286277 (Scopus ID)
Note

QC 20260603

Available from: 2026-06-03 Created: 2026-06-03 Last updated: 2026-06-03Bibliographically approved
Agram, N., Espen Benth, F., Pucci, G. & Rems, J. (2025). A Deep Learning Approach to Renewable Capacity Installation under Jump Uncertainty.
Open this publication in new window or tab >>A Deep Learning Approach to Renewable Capacity Installation under Jump Uncertainty
2025 (English)Manuscript (preprint) (Other academic)
Abstract [en]

We study a stochastic model for the installation of renewable energy capacity under demand uncertainty and jump driven dynamics. The system is governed by a multidimensional Ornstein-Uhlenbeck (OU) process driven by a subordinator, capturing abrupt variations in renewable generation and electricity load. Installation decisions are modeled through control actions that increase capacity in response to environmental and economic conditions. We consider two distinct solution approaches. First, we implement a structured threshold based control rule, where capacity is increased proportionally when the stochastic capacity factor falls below a fixed level. This formulation leads to a nonlinear partial integro-differential equation (PIDE), which we solve by reformulating it as a backward stochastic differential equation with jumps. We extend the DBDP solver in [15] to the pure jump setting, employing a dual neural network architecture to approximate both the value function and the jump sensitivity. Second, we propose a fully data driven deep control algorithm that directly learns the optimal feedback policy by minimizing the expected cost functional using neural networks. This approach avoids assumptions on the form of the control rule and enables adaptive interventions based on the evolving system state. Numerical experiments highlight the strengths of both methods. While the threshold based BSDE approach offers interpretability and tractability, the deep control strategy achieves improved performance through flexibility in capacity allocation. Together, these tools provide a robust framework for decision support in long term renewable energy expansion under uncertainty

Keywords
Capacity installation; Renewable energy expansion; Jump diffusion; BSDEs; Deep learning; Optimal stochastic control; Threshold control.
National Category
Computational Mathematics
Identifiers
urn:nbn:se:kth:diva-374875 (URN)10.48550/arXiv.2509.12364 (DOI)
Note

QC 20260107

Available from: 2026-01-06 Created: 2026-01-06 Last updated: 2026-01-12Bibliographically approved
Agram, N. & Pucci, G. (2025). Deep BSVIEs Parametrization and Learning-Based Applications.
Open this publication in new window or tab >>Deep BSVIEs Parametrization and Learning-Based Applications
2025 (English)In: Article in journal (Refereed) Submitted
National Category
Computational Mathematics
Identifiers
urn:nbn:se:kth:diva-366385 (URN)
Note

QC 20250804

Available from: 2025-07-07 Created: 2025-07-07 Last updated: 2026-01-12Bibliographically approved
Agram, N., Arharas, I., Pucci, G. & Rems, J. (2025). Deep Learning for Energy Market Contracts: Dynkin Game with Doubly RBSDEs.
Open this publication in new window or tab >>Deep Learning for Energy Market Contracts: Dynkin Game with Doubly RBSDEs
2025 (English)Manuscript (preprint) (Other academic)
Abstract [en]

We formulate a Contract for Difference (CfD) with early exit options as a two-player zero-sum Dynkin game, reflecting the strategic interaction between an electricity producer and a regulatory entity. The game incorporates penalties for early termination and mean-reverting price dynamics, with the value characterized through a doubly reflected backward stochastic differential equation (DRBSDE). To compute the contract value and optimal stopping strategies, we develop a neural solver that approximates the DRBSDE solution using a sequence of neural networks trained on simulated trajectories. The method avoids discretizing the state space, supports time-dependent barriers, and scales to highdimensional settings. We establish a convergence result and test the method on two scenarios: a benchmark symmetric game in 20 dimensions, and a CfD model with 24-dimensional electricity prices representing multiple European zones. The results demonstrate that the proposed solver accurately captures the contract’s value and optimal stopping regions, with consistent performance across dimensional settings.

Keywords
Deep learning, Doubly reflected BSDEs, Contract for difference, Dynkin game
National Category
Computational Mathematics
Identifiers
urn:nbn:se:kth:diva-366383 (URN)10.48550/arXiv.2503.00880 (DOI)
Note

QC 20260604

Available from: 2025-07-07 Created: 2025-07-07 Last updated: 2026-06-04Bibliographically approved
Agram, N., Espen Benth, F. & Pucci, G. (2025). Installation of renewable capacities to meet energy demand and emission constraints under uncertainty. IMA Journal of Management Mathematics, 37(1), 39-60
Open this publication in new window or tab >>Installation of renewable capacities to meet energy demand and emission constraints under uncertainty
2025 (English)In: IMA Journal of Management Mathematics, ISSN 1471-678X, E-ISSN 1471-6798, Vol. 37, no 1, p. 39-60Article in journal (Refereed) Published
Abstract [en]

This paper focuses on minimizing the costs of installing renewable energy capacity while meeting emission constraints under uncertainty in both energy demand and renewable production. We consider a setting where decision-makers must determine when and how much renewable capacity to install, balancing investment costs with future emissions. Our optimization problem combines cost minimization with a probabilistic constraint on total accumulated emissions, reflecting regulatory limits that may be exceeded only with small probability. We examine different investment strategies, allowing for one or multiple installation times, and provide explicit solutions in simplified cases. Our main insight is that, under reasonable assumptions on costs and uncertainty, a single, well-timed investment is optimal and may be delayed to reduce costs when uncertainty and discounting are accounted for. These results challenge common stepwise installation strategies and suggest that committing to a single large investment, possibly postponed, may be more cost-effective and efficient in reaching emission targets. Our findings offer practical guidance for policymakers and energy planners on how to balance costs, timing and environmental goals when expanding renewable energy capacity under uncertainty.

Place, publisher, year, edition, pages
Oxford University Press (OUP), 2025
Keywords
optimization with probabilistic constraints, capacity expansion, energy systems, renewable energy, emission reduction
National Category
Engineering and Technology
Identifiers
urn:nbn:se:kth:diva-366371 (URN)10.1093/imaman/dpaf023 (DOI)001520418800001 ()2-s2.0-105027317659 (Scopus ID)
Funder
Swedish Research Council, 2020-04697
Note

QC 20260127

Available from: 2025-07-07 Created: 2025-07-07 Last updated: 2026-01-27Bibliographically approved
Agram, N., Pucci, G. & Øksendal, B. (2024). Impulse Control of Conditional McKean–Vlasov Jump Diffusions. Journal of Optimization Theory and Applications, 200(3), 1100-1130
Open this publication in new window or tab >>Impulse Control of Conditional McKean–Vlasov Jump Diffusions
2024 (English)In: Journal of Optimization Theory and Applications, ISSN 0022-3239, E-ISSN 1573-2878, Vol. 200, no 3, p. 1100-1130Article in journal (Refereed) Published
Abstract [en]

In this paper, we consider impulse control problems involving conditional McKean–Vlasov jump diffusions, with the common noise coming from the σ-algebra generated by the first components of a Brownian motion and an independent compensated Poisson random measure. We first study the well-posedness of the conditional McKean–Vlasov stochastic differential equations (SDEs) with jumps. Then, we prove the associated Fokker–Planck stochastic partial differential equation (SPDE) with jumps. Next, we establish a verification theorem for impulse control problems involving conditional McKean–Vlasov jump diffusions. We obtain a Markovian system by combining the state equation with the associated Fokker–Planck SPDE for the conditional law of the state. Then we derive sufficient variational inequalities for a function to be the value function of the impulse control problem, and for an impulse control to be the optimal control. We illustrate our results by applying them to the study of an optimal stream of dividends under transaction costs. We obtain the solution explicitly by finding a function and an associated impulse control, which satisfy the verification theorem.

Place, publisher, year, edition, pages
Springer Nature, 2024
National Category
Mathematics
Identifiers
urn:nbn:se:kth:diva-346045 (URN)10.1007/s10957-023-02370-6 (DOI)001144857000001 ()2-s2.0-85182649962 (Scopus ID)
Funder
Swedish Research Council, 2020-04697KTH Royal Institute of Technology
Note

QC 20240502

Available from: 2024-05-01 Created: 2026-01-01 Last updated: 2026-01-12Bibliographically approved
Organisations
Identifiers
ORCID iD: ORCID iD iconorcid.org/0000-0001-9065-1410

Search in DiVA

Show all publications