MétaCan
Menu
Back to cohort
Record W6992834543

Modeling and State Estimation of Bio-processes using Dynamic Flux Balances

2023· dissertation· en· W6992834543 on OpenAlexfundno aff

Bibliographic record

VenueUWSpace (University of Waterloo) · 2023
Typedissertation
Languageen
FieldBiochemistry, Genetics and Molecular Biology
TopicMicrobial Metabolic Engineering and Bioproduction
Canadian institutionsnot available
FundersNatural Sciences and Engineering Research Council of CanadaMitacsSanofi
KeywordsProcess (computing)Work (physics)Context (archaeology)LimitingProteogenomics
DOInot available

Abstract

fetched live from OpenAlex

Due to the increasing demand for bio-pharmaceuticals, optimization of bio-processes' productivity and reduction of process variability have become critical goals for manufacturers. Mathematical models of the fermentation processes are instrumental in achieving these goals. \n \nDynamic flux balance analysis (DFBA), sometimes also referred to as dynamic flux balance modeling (DFBM), is a type of mechanistic modeling approach that can describe the dynamic evolution of key metabolites based on the structure of metabolic networks. DFBA predicts the dynamic evolution of metabolites based on the assumption that resources are optimally allocated so as to maximize/minimize a biological objective function, e.g. maximization of cell growth. Accordingly, DFBA is formulated by a linear programming (LP) problem to compute the metabolic fluxes at each time interval. Then, the evolution of concentrations of different metabolites over time is obtained from the integration of mass balances that are based on the calculated fluxes. \n \nGenerally, the LP used to solve a DFBM for a particular microorganism may have multiple solutions. Mathematically, the multiplicity of solutions arises due to the under-determinancy of the LP. On the other hand, from the biological point of view, the occurrence of multiple solutions may correctly describe the behavior of different strains of the same microorganism or alternatively the occurrence of metabolism switches under different operating conditions. The choice of one solution in the presence of multiplicity is further complicated by the fact that different commercial solvers may lead to different solutions of identical LPs. However, a good DFBA model should be solver-independent while it should be able to correctly describe available data for a specific microorganism strain. \n \nFollowing the above a good LP solver should choose the specific solution based on the strain instead of choosing the solution "randomly" as most commercial solvers do. Hence, the first contribution of this research is to construct a solver that can select a specific solution among all possible optima that is compatible with experimental data. The weighted primal-dual method (WPDM) presented in Chapter 3, is a modified version of the interior point method (IPM) which uses interior weights to solve the LP. By manipulating these weights, the specific optimal solution can be obtained when multiple optimal solutions occur. The interior weights can be found by fitting experimental data obtained for a specific strain of a microorganism. \n \nAlthough WPDM was able to select optimal solutions to fit the data, it was found to be computationally expensive and thus less suitable for large networks. To address this, an alternative fast and low-code algorithm called the ellipsoidal reflection method (ERM) was developed as described in Chapter 6. This algorithm is able to select particular solutions among all possible solutions based on the combination of quadratic programming (QP) and LP problems. ERM plays the same role in DFBM but it can greatly reduce the computations thus making it suitable for future real-time applications. \n \nAn important application of mechanistic models such as DFBM in bioreactors is for the purpose of estimation of states that cannot be measured directly from available measurements. The ability of estimate variables such as growth rate, productivity or key nutrients are crucial for controlling and optimizing the process. State estimation for biochemical systems is particularly difficult due to the lack of online measurements in industrial bio-processes. While variables such as dissolved oxygen, temperature and pH are regularly measured and controlled, most metabolites' concentrations cannot be measured online. Thus, lack of observability of unmeasured states from measured ones are a known challenge in bio-processes. \n \nTo address the lack of observability, set membership estimation (SME) is proposed whereby the upper and lower bounds of each state are estimated based on limited measurements. This approach is motivated by the fact that the cell culture media recipe is generally fixed and the variations of the initial concentrations with respect to the nominal recipe are within small ranges. The SME treats the variation of initial concentrations as a set and propagates the initial bounds of the set onto the bounds of each metabolite at each time step. In this research, two methods of SME are proposed to estimate the bounds of metabolites. \n \nThe first state estimation method, described in chapter 4, is based on the identification of active constraints and assumes that the solution is always unique in DFBA. Since the concentration is varying with time, the LP problem in DFBA can be formulated as an LP with varying parameters. Then, Multiparametric linear programming (mpLP) can be used to convert the DFBA system into a variable structure system (VSS). VSS describes the system as composed of multiple subsystems where each subsystem describes a different region of the state space. For each subsystem, an extended Kalman filter (EKF) is constructed to estimate the key states, and the remaining states are estimated by SME. Moreover, the states crossing in or out of each region of the state space are monitored by a special algorithm and switches between different EKFs are determined accordingly. In the \\textit{E. coli} model, it was assumed that only biomass and culture volume are measured and are used to estimate the bounds of the other states. \n \nThe second state estimation method presented in chapter 5 is an extension of the first method but it explicitly considers the existence of multiple solutions. In this second method, WPDM is used to replace the LP solver in DFBA and multiparametric nonlinear programming (mpNLP) is employed to solve the WPDM interior point-based algorithm. To propagate the uncertain sets by nonlinear mapping, the sets are split into smaller sets and are propagated separately by a linear mapping approximation. This is followed by an assembly operation of all these mapped sets together into one set for each state. Again, for the E. coli model, only biomass and culture volume are assumed to be measured and are used to estimate bounds on the other states. This method is shown to generate bounds of all states much faster than a Monte Carlo algorithm. \n \nTo test these methods proposed a platform of culturing B. pertussis has been set up. In chapter 7, a batch culture of B. pertussis and modeling by DFBM are presented. The protocols of shake flask, batch culture, and measurements of concentrations of amino acids in the culture by HPLC are set up. To solve the multiplicity issue, ERM is used in the modeling by DFBM. Based on the experimental data, DFBM adapted from the previous model is used to fit. The DFBM model can roughly capture the dynamics of key amino acids but not of all of them.

Fetched live from OpenAlex and de-inverted. Abstracts are not stored in this database: the inverted indexes are 8.6 GB of the frame’s 9.3 GB of text, and the host has 13 GB free.

How this classification was reachedexpand

Full frame machine prediction

Teacher imitation

Not calibrated prevalence, not ground truth. Human validation pending. The Gemma side is a direct model label for every work in the frame, read from the title-only record. The Codex side is a classifier learned from the 10,348 direct Codex labels and calibrated to design-weighted sample rates; fields without enough sample support carry no Codex call. Candidate is the union of the two sides; consensus is their intersection. These outputs are machine_predicted_unvalidated and are not human labels.

metaresearch head score (Codex)0.001
metaresearch head score (Gemma)0.002
Version: metacan-v3-hybrid-931329e0061cValidation status: machine_predicted_unvalidated
Candidate categoriesnone
Consensus categoriesnone
DomainCandidate signal: none · Consensus signal: none
Study designCandidate signal: Simulation or modeling · Consensus signal: Simulation or modeling
GenreCandidate signal: Empirical · Consensus signal: none
Teacher disagreement score0.011
Threshold uncertainty score0.022

Distilled classifier scores by category (both heads)

CategoryCodexGemma
Metaresearch0.0010.002
Meta-epidemiology (narrow)0.0010.001
Meta-epidemiology (broad)0.0010.001
Bibliometrics0.0010.001
Science and technology studies0.0010.001
Scholarly communication0.0020.001
Open science0.0010.001
Research integrity0.0020.001
Insufficient payload (model declined to judge)0.0020.000

Machine scores (provisional)

The two teacher heads of the student model, read on this work. A score orders the frame for review; it never asserts a category, and the validation status ships verbatim with every row.

Baseline scores from an immature model (maturity gate not passed, 7 training rounds). Scores rank; they never assert a category.

Opus teacher head0.007
GPT teacher head0.210
Teacher spread0.203 · how far apart the two teachers sit on this one work
Validation statusscore_only:v0-immature-baseline · verbatim from the scoring run: score_only means the number may rank works, and no category label ships from it

Classification

machine, unvalidated

Machine predicted; a candidate call from one source (direct Gemma or distilled Codex), not a consensus.

The models applied no category: nothing in the taxonomy fit this work.
Study designSimulation or modeling
Domainnot available
GenreEmpirical

How this classification was reached, model by model and score by score, is at the end of the page under "How this classification was reached".

Quick stats

Citations0
Published2023
Admission routes1
Has abstractyes

Explore more

Same venueUWSpace (University of Waterloo)Same topicMicrobial Metabolic Engineering and BioproductionFrench-language works237,207