{"id":489110,"date":"2026-07-30T05:11:12","date_gmt":"2026-07-30T05:11:12","guid":{"rendered":"https:\/\/savepearlharbor.com\/?p=489110"},"modified":"-0001-11-30T00:00:00","modified_gmt":"-0001-11-29T21:00:00","slug":"","status":"publish","type":"post","link":"https:\/\/savepearlharbor.com\/?p=489110","title":{"rendered":"Intelligent systems at phystech: 2026 graduation"},"content":{"rendered":"<div xmlns=\"http:\/\/www.w3.org\/1999\/xhtml\">\n<figure class=\"full-width \"><img decoding=\"async\" src=\"https:\/\/habrastorage.org\/r\/w1560\/getpro\/habr\/upload_files\/31d\/b61\/aa7\/31db61aa7110b59d43d1fcb5ecaded33.png\" width=\"1018\" height=\"812\" sizes=\"auto, (max-width: 780px) 100vw, 50vw\" srcset=\"https:\/\/habrastorage.org\/r\/w780\/getpro\/habr\/upload_files\/31d\/b61\/aa7\/31db61aa7110b59d43d1fcb5ecaded33.png 780w,&#10;       https:\/\/habrastorage.org\/r\/w1560\/getpro\/habr\/upload_files\/31d\/b61\/aa7\/31db61aa7110b59d43d1fcb5ecaded33.png 781w\" loading=\"lazy\" decode=\"async\"\/><\/figure>\n<p>Literature review, comparison, and analysis of existing solutions constitute a significant part of a thesis manuscript.\u00a0 In the last three years, AI chats have dramatically intervened in the thesis preparation workflow. The ability of AI chats to analyze and compare discourages students from acting as the primary drivers of research and instead encourages their passive reliance on AI-generated content. This situation makes manuscripts eclectic and lacks focus. The theoretical part disconnects from the computational experiments. As a result, the audience is reluctant to read students&#8217; manuscripts. Ironically, they use AI chats to summarise the manuscripts to extract their main contributions. We are proud to say that our thesis works do not suffer from this AI distortion. We are grateful to our students, especially our bachelor\u2019s and master\u2019s alumni, who have carefully prepared and successfully defended their theses, producing clear, well-founded, and meaningful results. Our thesis manuscripts are designed for public discussion and rigorous scientific analysis. Each thesis presents a clear research idea grounded in the fundamentals of machine learning. It includes a theoretical justification, a repository with the source code of the computational experiments, and slides. Videos of the pre-defence sessions are uploaded to our <a href=\"https:\/\/www.youtube.com\/c\/MachineLearningPhystech\" rel=\"noopener noreferrer nofollow\">YouTube<\/a> channel \u201cIntelligent Systems\u201d:\u00a0<a href=\"https:\/\/www.youtube.com\/watch?v=9PvcQsX9TK0&amp;t=16s\" rel=\"noopener noreferrer nofollow\">Bachelors<\/a> &amp; <a href=\"https:\/\/www.youtube.com\/watch?v=LzIn6i5FB2Y\" rel=\"noopener noreferrer nofollow\">Masters<\/a>.<\/p>\n<h2>Applied methods in machine learning<\/h2>\n<p>Several of this year&#8217;s theses take on hidden dynamics and structure in real-world data \u2014 from feedback effects inside deployed ML systems to signals recorded from the human brain and body.<\/p>\n<p><strong>Andrey Veprikov<\/strong>&#8216;s <a href=\"https:\/\/github.com\/intsystems\/Veprikov-MS-Thesis\/blob\/main\/paper\/main_veprikov.pdf\" rel=\"noopener noreferrer nofollow\">thesis<\/a>, supervised by <strong>Dr. Anton Khritankov, <\/strong>addresses machine learning systems in which the model&#8217;s own predictions quietly change the data it will see next. This hidden feedback loop often causes models to degrade over time even as the loss on the current batch keeps improving. He introduces a single control parameter into the training step and proves conditions under which it keeps the system stable, validating the approach both on synthetic tasks and on a real medical-imaging benchmark. Parts of the work appeared in <em>Knowledge and Information Systems<\/em> (Springer), in <em>Artificial Intelligence and Decision Making<\/em>, and at the MMRO conference.<\/p>\n<p><strong>Sergey Dementev<\/strong>&#8216;s <a href=\"https:\/\/github.com\/sdem3\/Hidden-Feedback-Loops-At-RecSys\/blob\/main\/paper\/main.pdf\" rel=\"noopener noreferrer nofollow\">research<\/a>, also supervised by Dr. Anton Khritankov,\u00a0 studies the same class of hidden feedback loops in recommender systems, where models repeatedly learn from data shaped by their own past recommendations. He shows that personalization degradation is not a single effect but arises through two distinct mechanisms: the contraction of user-preference diversity and the retraining-driven reshaping of served content. The work provides mathematical guarantees for these effects, including activation conditions and a measurable threshold for personalization collapse. Experiments on synthetic data and MovieLens-20M confirm the theory and show that explicit diversification can preserve diversity without sacrificing recommendation quality.<\/p>\n<figure class=\"full-width \"><img decoding=\"async\" src=\"https:\/\/habrastorage.org\/r\/w1560\/getpro\/habr\/upload_files\/a8b\/155\/e63\/a8b155e63ffacc5a8aa4b62d1573515b.png\" alt=\"The typical process of obtaining recommendations by recommendation systems.\" title=\"The typical process of obtaining recommendations by recommendation systems.\" width=\"964\" height=\"320\" sizes=\"auto, (max-width: 780px) 100vw, 50vw\" srcset=\"https:\/\/habrastorage.org\/r\/w780\/getpro\/habr\/upload_files\/a8b\/155\/e63\/a8b155e63ffacc5a8aa4b62d1573515b.png 780w,&#10;       https:\/\/habrastorage.org\/r\/w1560\/getpro\/habr\/upload_files\/a8b\/155\/e63\/a8b155e63ffacc5a8aa4b62d1573515b.png 781w\" loading=\"lazy\" decode=\"async\"\/><\/p>\n<div><figcaption>The typical process of obtaining recommendations by recommendation systems.<\/figcaption><\/div>\n<\/figure>\n<p><strong>Alexander Terentyev<\/strong>&#8216;s <a href=\"https:\/\/github.com\/intsystems\/Terentev-BS-Thesis\/tree\/master\/slides\/Slides.pdf\" rel=\"noopener noreferrer nofollow\">thesis<\/a>, supervised by our alumnus and professor <strong>Dr. Roman Isachenko<\/strong>, studies an inverse problem for partial differential equations. It introduces a regularization method that produces physically meaningful solutions and allows control over their spatial structure, tested on an EEG dataset. Unlike other approaches, it does not impose strict assumptions on the properties of the sources \u2014 their structure follows directly from the regularization.<\/p>\n<div class=\"floating-image\">\n<figure class=\"float full-width \"><img decoding=\"async\" src=\"https:\/\/habrastorage.org\/r\/w1560\/getpro\/habr\/upload_files\/233\/88c\/f58\/23388cf58d7acc9cb76007d2c9b522e2.png\" alt=\"Examples of CT image translation from contrast (bottom row) to native (non-contrast) in three projections (second from the bottom) and a heatmap of the difference between the translated image and the true native image (right side).\" title=\"Examples of CT image translation from contrast (bottom row) to native (non-contrast) in three projections (second from the bottom) and a heatmap of the difference between the translated image and the true native image (right side).\" width=\"634\" height=\"520\" sizes=\"auto, (max-width: 780px) 100vw, 50vw\" srcset=\"https:\/\/habrastorage.org\/r\/w780\/getpro\/habr\/upload_files\/233\/88c\/f58\/23388cf58d7acc9cb76007d2c9b522e2.png 780w,&#10;       https:\/\/habrastorage.org\/r\/w1560\/getpro\/habr\/upload_files\/233\/88c\/f58\/23388cf58d7acc9cb76007d2c9b522e2.png 781w\" loading=\"lazy\" decode=\"async\"\/><\/p>\n<div><figcaption>Examples of CT image translation from contrast (bottom row) to native (non-contrast) in three projections (second from the bottom) and a heatmap of the difference between the translated image and the true native image (right side).<\/figcaption><\/div>\n<\/figure>\n<p><strong>Matvei Kreinin<\/strong>&#8216;s <a href=\"https:\/\/github.com\/intsystems\/Kreinin-BS-Thesis\/blob\/master\/docs\/Kreinin_BS_Thesis.pdf\" rel=\"noopener noreferrer nofollow\">work<\/a>\u00a0 under supervision of <strong>Dr. Aleksandr Beznosikov, D.Sc.<\/strong>, addresses bidirectional translation between native (non-contrast) and contrast-enhanced abdominal computed tomography (CT) series by means of conditional flow matching theory. He proposes the Medical Flow Matching (MFM) framework together with a neural-network parameterization of the velocity field, TimeResNet. The method enables single-step deterministic inference via integration of an ordinary differential equation, and a single network realizes both translation directions.<\/p>\n<figure class=\"float full-width \"><img decoding=\"async\" src=\"https:\/\/habrastorage.org\/r\/w1560\/getpro\/habr\/upload_files\/17f\/c40\/0a0\/17fc400a067c63b68fcbae4d4200db8b.png\" width=\"968\" height=\"582\" sizes=\"auto, (max-width: 780px) 100vw, 50vw\" srcset=\"https:\/\/habrastorage.org\/r\/w780\/getpro\/habr\/upload_files\/17f\/c40\/0a0\/17fc400a067c63b68fcbae4d4200db8b.png 780w,&#10;       https:\/\/habrastorage.org\/r\/w1560\/getpro\/habr\/upload_files\/17f\/c40\/0a0\/17fc400a067c63b68fcbae4d4200db8b.png 781w\" loading=\"lazy\" decode=\"async\"\/><\/figure>\n<p><strong>Daniil Dorin<\/strong>&#8216;s <a href=\"https:\/\/github.com\/intsystems\/Dorin-BS-Thesis\/blob\/master\/paper\/paper.pdf\" rel=\"noopener noreferrer nofollow\">thesis<\/a>, supervised by our alumnus and professor <strong>Dr. Andrey Grabovoy<\/strong>, is devoted to decoding and reconstructing visual stimuli from brain activity recorded simultaneously with fMRI and EEG \u2014 two neuroimaging modalities whose spatial and temporal resolutions complement each other. He frames decoding as a single problem approached in three stages of increasing structural complexity: estimating the hemodynamic delay of the BOLD signal, extracting spatiotemporal features of fMRI through Riemannian geometry and class-specific activity masks that stay robust under limited sample sizes, and finally a multimodal architecture that contrastively aligns joint fMRI\u2013EEG embeddings with the CLIP space and reconstructs the stimulus via a two-stage diffusion generation. Experiments show that combining the two modalities outperforms single-modality decoding, and that adding a diffusion prior further improves the semantic fidelity of the reconstructions. The thesis is based on 3 papers published during Daniil\u2019s study. The three stages of the work, ordered by the structure each one is able to represent:<a href=\"https:\/\/doi.org\/10.1007\/s13755-024-00315-5\" rel=\"noopener noreferrer nofollow\"> stage 1 models time alone<\/a>, recovering the hemodynamic delay \u0394t from a linear autoregressive forecast of fMRI given the stimulus; <a href=\"https:\/\/doi.org\/10.1134\/S1064562425700383\" rel=\"noopener noreferrer nofollow\">stage 2 adds space<\/a>, describing class-specific activity masks through the Riemannian geometry of their covariance matrices, which keeps the feature dimension independent of sample size; <a href=\"https:\/\/doi.org\/10.14357\/19922264260203\" rel=\"noopener noreferrer nofollow\">stage 3 adds a second modality<\/a>, contrastively aligning the joint fMRI\u2013EEG embedding with CLIP space before a two-stage diffusion reconstruction. The stages are sequential rather than parallel \u2014 the delay estimated in the first is what temporally aligns the two modalities in the third.<\/p>\n<p><strong>Anastasia German<\/strong>&#8216;s <a href=\"https:\/\/github.com\/AnastasiaGerman01\/2025-Project-172\/blob\/main\/paper\/%D0%B4%D0%B8%D0%BF%D0%BF%D0%BB%D0%BE%D0%BC%20%D0%B8%D1%82%D0%BE%D0%B3%D0%BE%D0%B2%D1%8B%D0%B8%CC%86.pdf\" rel=\"noopener noreferrer nofollow\">work<\/a>, also supervised by <strong>Dr. Andrey Grabovoy<\/strong>, addresses the prediction of fMRI images \u2014 more precisely, increments of the BOLD signal \u2014 from the audio and video streams of a naturalistic audiovisual film, using simple, interpretable voxel-wise linear models. Three models of increasing complexity are built: a unimodal audio baseline, based on MFCC features, a multimodal banded-ridge regression with separate penalties for audio and video, and \u2014 the central model \u2014 a voxel-wise gated model in which predictions from the two modalities are mixed by a scalar gate, whose optimal value is derived analytically in closed form. The key result: banded-ridge regression yields only a marginal improvement, while the gated model consistently outperforms both unimodal baselines across all nine subjects, confirmed statistically. The analysis also shows that the gate favors audio almost everywhere, that the contribution of video is weak and distributed, and that the gain is achieved specifically through local, voxel-wise modality selection unavailable to a globally fitted joint model.<\/p>\n<\/div>\n<div class=\"floating-image\">\n<figure class=\"float full-width \"><img decoding=\"async\" src=\"https:\/\/habrastorage.org\/r\/w1560\/getpro\/habr\/upload_files\/f88\/e26\/417\/f88e26417121396f652eece8db0fb4e3.png\" alt=\"Observed time series generated from linear ODE system and smoothed phase trajectories from multiple ODE systems.\" title=\"Observed time series generated from linear ODE system and smoothed phase trajectories from multiple ODE systems.\" width=\"988\" height=\"744\" sizes=\"auto, (max-width: 780px) 100vw, 50vw\" srcset=\"https:\/\/habrastorage.org\/r\/w780\/getpro\/habr\/upload_files\/f88\/e26\/417\/f88e26417121396f652eece8db0fb4e3.png 780w,&#10;       https:\/\/habrastorage.org\/r\/w1560\/getpro\/habr\/upload_files\/f88\/e26\/417\/f88e26417121396f652eece8db0fb4e3.png 781w\" loading=\"lazy\" decode=\"async\"\/><\/p>\n<div><figcaption>Observed time series generated from linear ODE system and smoothed phase trajectories from multiple ODE systems.<\/figcaption><\/div>\n<\/figure>\n<p><strong>Kirill Semkin<\/strong> in his <a href=\"https:\/\/github.com\/intsystems\/tssa_method\/blob\/master\/doc\/Tensor_SSA_Semkin.pdf\" rel=\"noopener noreferrer nofollow\">thesis<\/a>, supervised by <strong>Dr. Roman Isachenko<\/strong>, proposes a method for time series classification under the assumption that series from different classes are generated by distinct dynamical systems, modeled as ordinary or stochastic differential equations whose noisy phase trajectories form the observed data. The classifier assigns each series to the class whose dynamics best fit it in a certain sense. Since exact trajectories are rarely available, they are replaced with conditional expectations obtained by solving a smoothing problem. The method is validated on synthetic linear systems, on human activity recognition from inertial sensors, and on heart failure detection from ECG recordings.<\/p>\n<p><strong>Anastasiia Vozniuk\u2019s<\/strong> <a href=\"https:\/\/github.com\/intsystems\/TextQualityAnalyser\/tree\/main\" rel=\"noopener noreferrer nofollow\">thesis<\/a> under supervision of <strong>Dr. Andrey Grabovoy<\/strong> focuses on possibilities of analysing text coherence and cohesion based on graph structure of the text. It operates on building a tree of the logical structure of the text with Rhetorical Structure Theory and then calculating various values of semantic and structural coherence, as well as entity cohesion. At the end, one can obtain a unified measure of text cohesion that is flexible and interpretable.<\/p>\n<\/div>\n<h2>Optimization<\/h2>\n<p>This year&#8217;s optimization theses range from projection-free first-order methods to budget-constrained bandit learning and new matrix optimizers.<\/p>\n<p><strong>Igor Ignashin<\/strong>&#8216;s <a href=\"https:\/\/github.com\/intsystems\/Ignashin_MS_thesis\/blob\/master\/paper\/MasterThesis.pdf\" rel=\"noopener noreferrer nofollow\">thesis<\/a> under supervision of <strong>Dr. Andrey Grabovoy<\/strong> is devoted to projection-free Frank\u2013Wolfe methods for machine learning and large-scale network optimization. It develops and analyzes conjugate, stochastic block-coordinate, and decentralized block-coordinate variants, focusing on reducing the computational cost per iteration while preserving O(1\/t)-type convergence guarantees. The work combines theoretical convergence proofs with experiments on machine-learning tasks and traffic-assignment networks. The traffic-assignment part was developed into the paper in the peer-reviewed <a href=\"https:\/\/link.springer.com\/article\/10.1007\/s10958-025-08157-6\" rel=\"noopener noreferrer nofollow\">journal<\/a>.<\/p>\n<p><strong>Ilgam Latypov<\/strong>&#8216;s <a href=\"https:\/\/github.com\/intsystems\/Latypov-MS-Thesis\/blob\/main\/paper\/paper.pdf\" rel=\"noopener noreferrer nofollow\">work<\/a>, supervised by <strong>Dr. Yuriy Dorn<\/strong>, addresses the problem of allocating a limited training budget across multiple adaptive online-learning experts while still identifying the best-performing one. He proposes M-LCB, a UCB-style meta-algorithm that selectively trains the most promising experts and achieves near-optimal regret guarantees in the stochastic setting. Theoretical results are complemented by experiments showing that adaptive budget allocation substantially accelerates convergence to the optimal expert compared to existing approaches. The full paper is available on <a href=\"https:\/\/arxiv.org\/pdf\/2510.22654\" rel=\"noopener noreferrer nofollow\">arXiv<\/a>.<\/p>\n<figure class=\"full-width \"><img decoding=\"async\" src=\"https:\/\/habrastorage.org\/r\/w1560\/getpro\/habr\/upload_files\/279\/8e0\/103\/2798e01030886607a793d06ba1911ea8.png\" alt=\"F-Muon and S-Muon perform on par with Muon in large-scale transformer pretraining, despite having completely different geometry.\" title=\"F-Muon and S-Muon perform on par with Muon in large-scale transformer pretraining, despite having completely different geometry.\" width=\"984\" height=\"448\" sizes=\"auto, (max-width: 780px) 100vw, 50vw\" srcset=\"https:\/\/habrastorage.org\/r\/w780\/getpro\/habr\/upload_files\/279\/8e0\/103\/2798e01030886607a793d06ba1911ea8.png 780w,&#10;       https:\/\/habrastorage.org\/r\/w1560\/getpro\/habr\/upload_files\/279\/8e0\/103\/2798e01030886607a793d06ba1911ea8.png 781w\" loading=\"lazy\" decode=\"async\"\/><\/p>\n<div><figcaption>F-Muon and S-Muon perform on par with Muon in large-scale transformer pretraining, despite having completely different geometry.<\/figcaption><\/div>\n<\/figure>\n<div class=\"floating-image\">\n<p><strong>Alexey Kravatskiy<\/strong>&#8216;s bachelor&#8217;s <a href=\"https:\/\/github.com\/intsystems\/Kravatskiy-BSc-Thesis\/blob\/main\/paper\/ky_fan_thesis.pdf\" rel=\"noopener noreferrer nofollow\">thesis<\/a>, supervised by <strong>Dr. Ivan Oseledets<\/strong>, <a href=\"http:\/\/D.Sc\" rel=\"noopener noreferrer nofollow\"><strong>D.Sc<\/strong><\/a><strong>.<\/strong>, investigates alternative norm choices within the framework of a norm-constrained linear minimization oracle. Novel matrix optimization algorithms are developed by generalizing the Muon optimizer using duals to Ky Fan norms and conic combinations of norms. It is demonstrated that competitive performance is not unique to Muon&#8217;s spectral norm and can be achieved with a broader family of norm constraints: the resulting F-Muon and S-Muon methods perform on par with Muon in large-scale transformer pretraining despite having completely different geometry. The thesis is based on a <a href=\"https:\/\/arxiv.org\/pdf\/2512.09678\" rel=\"noopener noreferrer nofollow\">paper<\/a> presented orally at ICOMP 2025, where it received the Best Paper award.<\/p>\n<\/div>\n<h2>Machine learning fundamentals and computational mathematics<\/h2>\n<p>The largest group of theses this year develops theory \u2014 new estimators, generalization bounds, and structural models for representation learning, architecture search, and generative modeling.<\/p>\n<p><strong>Iryna Zabarianska<\/strong>&#8216;s <a href=\"https:\/\/github.com\/intsystems\/DBMI-\/blob\/main\/thesis.pdf\" rel=\"noopener noreferrer nofollow\">work<\/a> under supervision of <strong>Dr. Alexander Korotin<\/strong>, addresses the estimation of mutual information (MI) between discrete random variables, which remains challenging for conventional methods as dimensionality or alphabet size grows. Leveraging recent advances in discrete bridge matching \u2014 a generative framework for domain transfer \u2014 she constructs a novel MI estimator, DBMI (Discrete Bridge Mutual Information). By expressing MI as a KL divergence between two reciprocal processes and using their Markov-chain representation, DBMI naturally handles discrete data. The method is evaluated on both a low-dimensional benchmark and a high-dimensional image benchmark with known ground truth, demonstrating superior performance over existing discrete MI estimators. The work is presented at the ICLR delta <a href=\"https:\/\/openreview.net\/challenge?redirect=%2Fforum%3Fid%3DyDYxBdDa5r\" rel=\"noopener noreferrer nofollow\">workshop<\/a>.<\/p>\n<p><strong>Maksim Ivanov<\/strong>&#8216;s bachelor&#8217;s <a href=\"https:\/\/github.com\/intsystems\/Zero-shot-structural-pruning\/blob\/master\/thesis\/thesis.pdf\" rel=\"noopener noreferrer nofollow\">thesis<\/a>, supervised by our alumnus and professor <strong>Dr. Oleg Bakhteev<\/strong>, proposes a method for structural pruning in neural networks, based on analyzing a deep-learning computation graph and estimating the information flow propagated through it. Compared with existing baselines, the proposed approach better preserves the monotonic relationship between the true and predicted losses.<\/p>\n<div class=\"floating-image\">\n<figure class=\"float full-width \"><img decoding=\"async\" src=\"https:\/\/habrastorage.org\/r\/w1560\/getpro\/habr\/upload_files\/0ab\/356\/65f\/0ab35665f8f00eefd0c89d5d17d68d30.png\" alt=\"An example of conversion of the computational graph of the target model into a surrogate graph for further edge pruning.\" title=\"An example of conversion of the computational graph of the target model into a surrogate graph for further edge pruning.\" width=\"728\" height=\"568\" sizes=\"auto, (max-width: 780px) 100vw, 50vw\" srcset=\"https:\/\/habrastorage.org\/r\/w780\/getpro\/habr\/upload_files\/0ab\/356\/65f\/0ab35665f8f00eefd0c89d5d17d68d30.png 780w,&#10;       https:\/\/habrastorage.org\/r\/w1560\/getpro\/habr\/upload_files\/0ab\/356\/65f\/0ab35665f8f00eefd0c89d5d17d68d30.png 781w\" loading=\"lazy\" decode=\"async\"\/><\/p>\n<div><figcaption>An example of conversion of the computational graph of the target model into a surrogate graph for further edge pruning.<\/figcaption><\/div>\n<\/figure>\n<p><strong>Nikita Okhotnikov<\/strong>&#8216;s master&#8217;s <a href=\"https:\/\/github.com\/intsystems\/lora_stable_rank\/blob\/main\/paper\/main.pdf\" rel=\"noopener noreferrer nofollow\">thesis<\/a>, supervised by <strong>Dr. Andrey Grabovoy<\/strong>, investigates the theoretical foundations of LoRA, one of the most widely used parameter-efficient fine-tuning techniques. The study builds on the previously observed implicit low-stable-rank bias of LoRA training and derives new expressivity and generalization bounds that incorporate this bias into theoretical estimates. Nikita argues that stable rank, rather than conventional algebraic rank, provides a more realistic measure of an adapter&#8217;s capacity and generalization ability.<\/p>\n<\/div>\n<div class=\"floating-image\">\n<figure class=\"float full-width \"><img decoding=\"async\" src=\"https:\/\/habrastorage.org\/r\/w1560\/getpro\/habr\/upload_files\/3aa\/69b\/eff\/3aa69beff9e99868724b6ca98d55e0b8.png\" alt=\"Curvature-aware subspace criterion, which restricts probing to the principal Hessian subspace spanned by the top curvature directions.\" title=\"Curvature-aware subspace criterion, which restricts probing to the principal Hessian subspace spanned by the top curvature directions.\" width=\"1030\" height=\"652\" sizes=\"auto, (max-width: 780px) 100vw, 50vw\" srcset=\"https:\/\/habrastorage.org\/r\/w780\/getpro\/habr\/upload_files\/3aa\/69b\/eff\/3aa69beff9e99868724b6ca98d55e0b8.png 780w,&#10;       https:\/\/habrastorage.org\/r\/w1560\/getpro\/habr\/upload_files\/3aa\/69b\/eff\/3aa69beff9e99868724b6ca98d55e0b8.png 781w\" loading=\"lazy\" decode=\"async\"\/><\/p>\n<div><figcaption>Curvature-aware subspace criterion, which restricts probing to the principal Hessian subspace spanned by the top curvature directions.<\/figcaption><\/div>\n<\/figure>\n<p><strong>Nikita Kiselev<\/strong>&#8216;s thesis <a href=\"https:\/\/github.com\/intsystems\/Kiselev-MS-Thesis\/blob\/main\/paper\/main.pdf\" rel=\"noopener noreferrer nofollow\">https:\/\/github.com\/intsystems\/Kiselev-MS-Thesis\/blob\/main\/paper\/main.pdf<\/a>, also supervised by <strong>Dr. Andrey Grabovoy<\/strong>, asks how much training data a parametric model actually needs, and answers this through the stabilization of the empirical loss surface: once one more object no longer reshapes the landscape near the minimum, the sample can be called sufficient. Existing quantitative criteria average the loss increment isotropically over the entire parameter space, which makes them impractical for large networks. He introduces a unified p-convergence criterion equipped with a preference function over the parameter space, recovering existing estimates as special cases and adding a new one: probing restricted to the subspace spanned by the leading Hessian eigenvectors. For this criterion he proves the same convergence guarantees as before, but now scaling with the subspace size rather than the full number of parameters, along with several efficient estimation algorithms that avoid computing the full Hessian. Experiments on MLP\/MNIST and on the nanochat d6 transformer confirm the theory, with the subspace-based estimator several orders of magnitude faster than direct Monte Carlo at no loss of accuracy. The results of his master\u2019s thesis are based on five papers published in peer-reviewed journals [<a href=\"https:\/\/link.springer.com\/article\/10.1134\/S1064562424601987\" rel=\"noopener noreferrer nofollow\">1<\/a>, <a href=\"https:\/\/link.springer.com\/article\/10.1007\/s10287-024-00528-9\" rel=\"noopener noreferrer nofollow\">2<\/a>, <a href=\"https:\/\/link.springer.com\/article\/10.1134\/S0965542524702002\" rel=\"noopener noreferrer nofollow\">3<\/a>, <a href=\"https:\/\/www.mathnet.ru\/php\/archive.phtml?wshow=paper&amp;jrnid=tisp&amp;paperid=1198\" rel=\"noopener noreferrer nofollow\">4<\/a>] and conference <a href=\"https:\/\/ieeexplore.ieee.org\/document\/10899113\" rel=\"noopener noreferrer nofollow\">proceedings<\/a>.<\/p>\n<\/div>\n<p><strong>Egor Petrov<\/strong>&#8216;s bachelor&#8217;s <a href=\"https:\/\/github.com\/modernTalker\/BSc-Thesis\/blob\/main\/thesis\/Thesis.pdf\" rel=\"noopener noreferrer nofollow\">thesis<\/a>, also supervised by <strong>Dr. Andrey Grabovoy<\/strong>, studies the loss landscape of Transformer architectures from a second-order theoretical perspective. The work derives exact Jacobian and Hessian matrices for all components of a full post-norm Transformer block \u2014 including self-attention, residual connections, LayerNorm, and the feed-forward network \u2014 by applying and extending tools from matrix calculus. The analytical formulas are verified against PyTorch autograd with accuracy up to numerical precision, while also providing faster computation for several derivative blocks. These results open the way to a more detailed theoretical analysis of Transformers, including their optimization stability, curvature-aware training methods, and scaling behavior.<\/p>\n<div class=\"floating-image\">\n<figure class=\"float full-width \"><img decoding=\"async\" src=\"https:\/\/habrastorage.org\/r\/w1560\/getpro\/habr\/upload_files\/dd9\/69c\/8ae\/dd969c8ae4b7d9cfc436254f8fd0edfd.png\" alt=\"The scheme of the proposed method: each expert is represented by a separate architecture in the architecture space. Its contribution into the final mixture-of-experts model is assigned based on the dataset clustering.\" title=\"The scheme of the proposed method: each expert is represented by a separate architecture in the architecture space. Its contribution into the final mixture-of-experts model is assigned based on the dataset clustering.\" width=\"1024\" height=\"660\" sizes=\"auto, (max-width: 780px) 100vw, 50vw\" srcset=\"https:\/\/habrastorage.org\/r\/w780\/getpro\/habr\/upload_files\/dd9\/69c\/8ae\/dd969c8ae4b7d9cfc436254f8fd0edfd.png 780w,&#10;       https:\/\/habrastorage.org\/r\/w1560\/getpro\/habr\/upload_files\/dd9\/69c\/8ae\/dd969c8ae4b7d9cfc436254f8fd0edfd.png 781w\" loading=\"lazy\" decode=\"async\"\/><\/p>\n<div><figcaption>The scheme of the proposed method: each expert is represented by a separate architecture in the architecture space. Its contribution into the final mixture-of-experts model is assigned based on the dataset clustering.<\/figcaption><\/div>\n<\/figure>\n<p><strong>Petr Babkin<\/strong>&#8216;s master&#8217;s <a href=\"https:\/\/github.com\/intsystems\/2023-Project-120\/blob\/master\/paper\/EdgeNES_diploma-5.pdf\" rel=\"noopener noreferrer nofollow\">thesis<\/a>, supervised by <strong>Dr. Oleg Bakhteev<\/strong>, studies neural architecture search for Mixture-of-Experts (MoE) models that accounts for the cluster structure of the data. Since changing an expert&#8217;s architecture changes its routing \u2014 and thus the data it trains on \u2014 the architectures and the data partition must be optimized jointly. The work introduces a surrogate function that takes an expert architecture and a binary vector encoding a subset of data clusters, and predicts the quality the architecture would attain on that subset without actual training. A Surrogate-assisted Generalized EM (SGEM) algorithm then uses this surrogate to jointly find the optimal data subsets for the experts and their architectures, with a proven convergence guarantee. On a CIFAR-100\/SVHN mixture, the method turns experts into domain specialists without any source labels, outperforming random-MoE and shared-architecture baselines.<\/p>\n<\/div>\n<div class=\"floating-image\">\n<figure class=\"float full-width \"><img decoding=\"async\" src=\"https:\/\/habrastorage.org\/r\/w1560\/getpro\/habr\/upload_files\/027\/8ad\/60f\/0278ad60fbf7fec75f6dbe4610e51449.png\" alt=\"The scheme of the proposed method: to take into account the diversity of the models we build an additional surrogate model that embeds the model architecture into latent space: the distance between the architectures approximates the disagreement between them.\" title=\"The scheme of the proposed method: to take into account the diversity of the models we build an additional surrogate model that embeds the model architecture into latent space: the distance between the architectures approximates the disagreement between them.\" width=\"1038\" height=\"656\" sizes=\"auto, (max-width: 780px) 100vw, 50vw\" srcset=\"https:\/\/habrastorage.org\/r\/w780\/getpro\/habr\/upload_files\/027\/8ad\/60f\/0278ad60fbf7fec75f6dbe4610e51449.png 780w,&#10;       https:\/\/habrastorage.org\/r\/w1560\/getpro\/habr\/upload_files\/027\/8ad\/60f\/0278ad60fbf7fec75f6dbe4610e51449.png 781w\" loading=\"lazy\" decode=\"async\"\/><\/p>\n<div><figcaption>The scheme of the proposed method: to take into account the diversity of the models we build an additional surrogate model that embeds the model architecture into latent space: the distance between the architectures approximates the disagreement between them.<\/figcaption><\/div>\n<\/figure>\n<p><strong>Alexander Udeneev<\/strong>&#8216;s <a href=\"https:\/\/github.com\/intsystems\/predicator-function-for-neural-networks\/blob\/master\/paper\/Udeneev2025Surrogate.pdf\" rel=\"noopener noreferrer nofollow\">research<\/a>, supervised by <strong>Dr. Oleg Bakhteev<\/strong>, &#171;Surrogate assisted diversity estimation in neural ensemble search,&#187; introduces a surrogate-assisted framework for Neural Ensemble Search (NES) that jointly optimizes predictive accuracy and architectural diversity to overcome the computational intractability of traditional ensemble methods. Two independent surrogate functions are employed: one predicts individual model accuracy via regression, while the other maps architectures into a latent space using triplet loss to estimate diversity based on predictive dissimilarity. A greedy selection algorithm then constructs ensembles by filtering candidates for high predicted accuracy and iteratively adding models that maximize geometric distance in the diversity latent space. Experimental results on FashionMNIST, CIFAR-10, and CIFAR-100 demonstrate that the approach achieves competitive or superior performance compared to Deep Ensembles and Random Search by effectively balancing individual model quality with collective diversity. Alexander&#8217;s bachelor&#8217;s thesis was later <a href=\"https:\/\/link.springer.com\/chapter\/10.1007\/978-3-032-30612-8_12\" rel=\"noopener noreferrer nofollow\">published<\/a> and presented at the AIAI 2026 international conference.<\/p>\n<\/div>\n<p><strong>Aleksandr Astakhov<\/strong>&#8216;s <a href=\"https:\/\/github.com\/AleksandrAstakhov\/dynamic-systems\/blob\/dev\/docs\/thesis.pdf\" rel=\"noopener noreferrer nofollow\">thesis<\/a> under supervision of our professor <strong>Dr. Vadim Strijov, D.Sc.,<\/strong> proposes a model for distributed dynamical systems whose connectivity operator switches among a finite set of matrices, motivated by the observation that any symmetric approximation incurs an irreducible error once interactions become directed. The model combines Takens-embedding-based dimensionality reduction with a set of K low-rank, asymmetric bilinear operators, one per regime. The resulting model improves forecasting accuracy over a symmetric correlation baseline.<\/p>\n<figure class=\"full-width \"><img decoding=\"async\" src=\"https:\/\/habrastorage.org\/r\/w1560\/getpro\/habr\/upload_files\/a6f\/78a\/117\/a6f78a117b7c7a0de5b30f64a3686283.png\" alt=\"\" title=\"\" width=\"1018\" height=\"214\" sizes=\"auto, (max-width: 780px) 100vw, 50vw\" srcset=\"https:\/\/habrastorage.org\/r\/w780\/getpro\/habr\/upload_files\/a6f\/78a\/117\/a6f78a117b7c7a0de5b30f64a3686283.png 780w,&#10;       https:\/\/habrastorage.org\/r\/w1560\/getpro\/habr\/upload_files\/a6f\/78a\/117\/a6f78a117b7c7a0de5b30f64a3686283.png 781w\" loading=\"lazy\" decode=\"async\"\/><\/p>\n<div><figcaption>Each data point is mapped by the probability-flow ODE of a pretrained diffusion model to its trajectory, pairwise distances between trajectories form the kernel matrix (block structure reflects semantic clusters), and its log-determinant gives a single scalar complexity score.<\/figcaption><\/div>\n<\/figure>\n<p><strong>Danila Chernousov<\/strong>&#8216;s bachelor&#8217;s <a href=\"https:\/\/github.com\/intsystems\/Chernousov-BSc-Thesis\/blob\/main\/paper\/report.pdf\" rel=\"noopener noreferrer nofollow\">thesis<\/a>, supervised by <strong>Dr. Andrey Grabovoy<\/strong>, proposes an unsupervised, label-free measure of dataset complexity, motivated by the fact that the size of a training set says little about how diverse and informative it actually is. A valid complexity measure is defined as a non-negative function that is invariant to the order of the objects, subadditive, and monotone, so that adding an object never decreases complexity. The main result is that any embedding of the objects into a Hilbert space produces such a measure, built from the embedding&#8217;s Gaussian kernel. The required semantic embedding is supplied by a pretrained diffusion model, which maps every object to its deterministic probability-flow ODE trajectory in the Hilbert space, so the diffusion trajectory itself serves as the embedding.<\/p>\n<div class=\"floating-image\">\n<figure class=\"float full-width \"><img decoding=\"async\" src=\"https:\/\/habrastorage.org\/r\/w1560\/getpro\/habr\/upload_files\/ef7\/f0a\/bcb\/ef7f0abcb7d229ce6534832d84421206.png\" alt=\"Relation between data space and model\u2019s data representation space\" title=\"Relation between data space and model\u2019s data representation space\" width=\"1020\" height=\"496\" sizes=\"auto, (max-width: 780px) 100vw, 50vw\" srcset=\"https:\/\/habrastorage.org\/r\/w780\/getpro\/habr\/upload_files\/ef7\/f0a\/bcb\/ef7f0abcb7d229ce6534832d84421206.png 780w,&#10;       https:\/\/habrastorage.org\/r\/w1560\/getpro\/habr\/upload_files\/ef7\/f0a\/bcb\/ef7f0abcb7d229ce6534832d84421206.png 781w\" loading=\"lazy\" decode=\"async\"\/><\/p>\n<div><figcaption>Relation between data space and model\u2019s data representation space<\/figcaption><\/div>\n<\/figure>\n<p><strong>Maria Nikitina<\/strong> <a href=\"https:\/\/github.com\/intsystems\/Vector-space-of-generative-models\/blob\/master\/paper\/%D0%94%D0%B8%D0%BF%D0%BB%D0%BE%D0%BC.pdf\" rel=\"noopener noreferrer nofollow\">work<\/a>, supervised by <strong>Dr. Oleg Bakhteev<\/strong> and co-supervised by our alumnus <strong>Anton Bishuk<\/strong>, examines the relationship between the matrices of an autoencoder&#8217;s parameters and the statistical properties of the data it is trained on. The parameters of a trained model are proposed as a dense vector representation of the corresponding data sample. The resulting vector space is shown to be invariant to symmetries of the parameter space, to preserve information about mixtures present in the original data distribution, and to yield vector sets that are linearly separable when the underlying models are trained on data from different distributions. The results of Maria\u2019s work were published in a peer-reviewed <a href=\"https:\/\/ubs.mtas.ru\/archive\/search_results_new.php?publication_id=23414\" rel=\"noopener noreferrer nofollow\">journal<\/a>.<\/p>\n<\/div>\n<p><strong>Daniil Kazachkov<\/strong>&#8216;s <a href=\"https:\/\/github.com\/intsystems\/Kazachkov-BSc-Diploma\/blob\/main\/docs\/report\/BSc_diploma.pdf\" rel=\"noopener noreferrer nofollow\">thesis<\/a> under supervision of <strong>Dr. Alexander Korotin<\/strong>, studies how to recover hidden pairwise dependencies between two sets of observations when the original matching within each group has been lost. The problem is formulated as probabilistic inference, and two approaches are developed: a categorical variational autoencoder with a Gumbel\u2013Sinkhorn latent permutation, and a global shift-mixture model trained with a modified EM algorithm. Experiments on synthetic periodic data show that both methods can recover latent pairs and the underlying data distribution. The neural model generalizes to new groups without re-optimization, while the explicit mixture model achieves higher accuracy when its assumptions match the data structure.<\/p>\n<h2>Conclusion<\/h2>\n<p>The thesis defence process starts long before the final presentation. At the beginning of the academic year, each student is already in contact with thesis supervisors and has a research topic. All thesis supervisors in the Intelligent Systems group hold a Ph.D. or D.Sc. degree in Physics and Mathematics and have an extensive publication record. During the academic year, each student presents their work several times: at the Department&#8217;s reporting seminars and at scientific conferences. Every student also takes part in teaching activities to improve their scientific communication and pedagogical skills.<\/p>\n<p>Links to the thesis manuscripts from <a href=\"https:\/\/intsystems.github.io\/materials\/thesis\/\" rel=\"noopener noreferrer nofollow\">recent years<\/a>.<\/p>\n<\/div>\n<p>\u0441\u0441\u044b\u043b\u043a\u0430 \u043d\u0430 \u043e\u0440\u0438\u0433\u0438\u043d\u0430\u043b \u0441\u0442\u0430\u0442\u044c\u0438 <a href=\"https:\/\/habr.com\/ru\/articles\/1064710\/\">https:\/\/habr.com\/ru\/articles\/1064710\/<\/a><\/p>\n","protected":false},"excerpt":{"rendered":"<p>Literature review, comparison, and analysis of existing solutions constitute a significant part of a thesis manuscript.\u00a0 In the last three years, AI chats have dramatically intervened in the thesis preparation workflow. The ability of AI chats to analyze and compare discourages students from acting as the primary drivers of research and instead encourages their passive reliance on AI-generated content. This situation makes manuscripts eclectic and lacks focus. The theoretical part disconnects from the computational experiments. As a result, the audience is reluctant to read students&#8217; manuscripts. Ironically, they use AI chats to summarise the manuscripts to extract their main contributions. We are proud to say that our thesis works do not suffer from this AI distortion. We are grateful to our students, especially our bachelor\u2019s and master\u2019s alumni, who have carefully prepared and successfully defended their theses, producing clear, well-founded, and meaningful results. Our thesis manuscripts are designed for public discussion and rigorous scientific analysis. Each thesis presents a clear research idea grounded in the fundamentals of machine learning. It includes a theoretical justification, a repository with the source code of the computational experiments, and slides. Videos of the pre-defence sessions are uploaded to our YouTube channel \u201cIntelligent Systems\u201d:\u00a0Bachelors &amp; Masters.Applied methods in machine learningSeveral of this year&#8217;s theses take on hidden dynamics and structure in real-world data \u2014 from feedback effects inside deployed ML systems to signals recorded from the human brain and body.Andrey Veprikov&#8217;s thesis, supervised by Dr. Anton Khritankov, addresses machine learning systems in which the model&#8217;s own predictions quietly change the data it will see next. This hidden feedback loop often causes models to degrade over time even as the loss on the current batch keeps improving. He introduces a single control parameter into the training step and proves conditions under which it keeps the system stable, validating the approach both on synthetic tasks and on a real medical-imaging benchmark. Parts of the work appeared in Knowledge and Information Systems (Springer), in Artificial Intelligence and Decision Making, and at the MMRO conference.Sergey Dementev&#8217;s research, also supervised by Dr. Anton Khritankov,\u00a0 studies the same class of hidden feedback loops in recommender systems, where models repeatedly learn from data shaped by their own past recommendations. He shows that personalization degradation is not a single effect but arises through two distinct mechanisms: the contraction of user-preference diversity and the retraining-driven reshaping of served content. The work provides mathematical guarantees for these effects, including activation conditions and a measurable threshold for personalization collapse. Experiments on synthetic data and MovieLens-20M confirm the theory and show that explicit diversification can preserve diversity without sacrificing recommendation quality.The typical process of obtaining recommendations by recommendation systems.Alexander Terentyev&#8217;s thesis, supervised by our alumnus and professor Dr. Roman Isachenko, studies an inverse problem for partial differential equations. It introduces a regularization method that produces physically meaningful solutions and allows control over their spatial structure, tested on an EEG dataset. Unlike other approaches, it does not impose strict assumptions on the properties of the sources \u2014 their structure follows directly from the regularization.Examples of CT image translation from contrast (bottom row) to native (non-contrast) in three projections (second from the bottom) and a heatmap of the difference between the translated image and the true native image (right side).Matvei Kreinin&#8217;s work\u00a0 under supervision of Dr. Aleksandr Beznosikov, D.Sc., addresses bidirectional translation between native (non-contrast) and contrast-enhanced abdominal computed tomography (CT) series by means of conditional flow matching theory. He proposes the Medical Flow Matching (MFM) framework together with a neural-network parameterization of the velocity field, TimeResNet. The method enables single-step deterministic inference via integration of an ordinary differential equation, and a single network realizes both translation directions.Daniil Dorin&#8217;s thesis, supervised by our alumnus and professor Dr. Andrey Grabovoy, is devoted to decoding and reconstructing visual stimuli from brain activity recorded simultaneously with fMRI and EEG \u2014 two neuroimaging modalities whose spatial and temporal resolutions complement each other. He frames decoding as a single problem approached in three stages of increasing structural complexity: estimating the hemodynamic delay of the BOLD signal, extracting spatiotemporal features of fMRI through Riemannian geometry and class-specific activity masks that stay robust under limited sample sizes, and finally a multimodal architecture that contrastively aligns joint fMRI\u2013EEG embeddings with the CLIP space and reconstructs the stimulus via a two-stage diffusion generation. Experiments show that combining the two modalities outperforms single-modality decoding, and that adding a diffusion prior further improves the semantic fidelity of the reconstructions. The thesis is based on 3 papers published during Daniil\u2019s study. The three stages of the work, ordered by the structure each one is able to represent: stage 1 models time alone, recovering the hemodynamic delay \u0394t from a linear autoregressive forecast of fMRI given the stimulus; stage 2 adds space, describing class-specific activity masks through the Riemannian geometry of their covariance matrices, which keeps the feature dimension independent of sample size; stage 3 adds a second modality, contrastively aligning the joint fMRI\u2013EEG embedding with CLIP space before a two-stage diffusion reconstruction. The stages are sequential rather than parallel \u2014 the delay estimated in the first is what temporally aligns the two modalities in the third.Anastasia German&#8217;s work, also supervised by Dr. Andrey Grabovoy, addresses the prediction of fMRI images \u2014 more precisely, increments of the BOLD signal \u2014 from the audio and video streams of a naturalistic audiovisual film, using simple, interpretable voxel-wise linear models. Three models of increasing complexity are built: a unimodal audio baseline, based on MFCC features, a multimodal banded-ridge regression with separate penalties for audio and video, and \u2014 the central model \u2014 a voxel-wise gated model in which predictions from the two modalities are mixed by a scalar gate, whose optimal value is derived analytically in closed form. The key result: banded-ridge regression yields only a marginal improvement, while the gated model consistently outperforms both unimodal baselines across all nine subjects, confirmed statistically. The analysis also shows that the gate favors audio almost everywhere, that the contribution of video is weak and distributed, and that the gain is achieved specifically through local, voxel-wise modality selection unavailable to a globally fitted joint model.Observed time series generated from linear ODE system and smoothed phase trajectories from multiple ODE systems.Kirill Semkin in his thesis, supervised by Dr. Roman Isachenko, proposes a method for time series classification under the assumption that series from different classes are generated by distinct dynamical systems, modeled as ordinary or stochastic differential equations whose noisy phase trajectories form the observed data. The classifier assigns each series to the class whose dynamics best fit it in a certain sense. Since exact trajectories are rarely available, they are replaced with conditional expectations obtained by solving a smoothing problem. The method is validated on synthetic linear systems, on human activity recognition from inertial sensors, and on heart failure detection from ECG recordings.Anastasiia Vozniuk\u2019s thesis under supervision of Dr. Andrey Grabovoy focuses on possibilities of analysing text coherence and cohesion based on graph structure of the text. It operates on building a tree of the logical structure of the text with Rhetorical Structure Theory and then calculating various values of semantic and structural coherence, as well as entity cohesion. At the end, one can obtain a unified measure of text cohesion that is flexible and interpretable.OptimizationThis year&#8217;s optimization theses range from projection-free first-order methods to budget-constrained bandit learning and new matrix optimizers.Igor Ignashin&#8217;s thesis under supervision of Dr. Andrey Grabovoy is devoted to projection-free Frank\u2013Wolfe methods for machine learning and large-scale network optimization. It develops and analyzes conjugate, stochastic block-coordinate, and decentralized block-coordinate variants, focusing on reducing the computational cost per iteration while preserving O(1\/t)-type convergence guarantees. The work combines theoretical convergence proofs with experiments on machine-learning tasks and traffic-assignment networks. The traffic-assignment part was developed into the paper in the peer-reviewed journal.Ilgam Latypov&#8217;s work, supervised by Dr. Yuriy Dorn, addresses the problem of allocating a limited training budget across multiple adaptive online-learning experts while still identifying the best-performing one. He proposes M-LCB, a UCB-style meta-algorithm that selectively trains the most promising experts and achieves near-optimal regret guarantees in the stochastic setting. Theoretical results are complemented by experiments showing that adaptive budget allocation substantially accelerates convergence to the optimal expert compared to existing approaches. The full paper is available on arXiv.F-Muon and S-Muon perform on par with Muon in large-scale transformer pretraining, despite having completely different geometry.Alexey&#8230;<\/p>\n","protected":false},"author":1,"featured_media":0,"comment_status":"closed","ping_status":"closed","sticky":false,"template":"","format":"standard","meta":{"footnotes":""},"categories":[],"tags":[],"class_list":["post-489110","post","type-post","status-publish","format-standard","hentry"],"_links":{"self":[{"href":"https:\/\/savepearlharbor.com\/index.php?rest_route=\/wp\/v2\/posts\/489110","targetHints":{"allow":["GET"]}}],"collection":[{"href":"https:\/\/savepearlharbor.com\/index.php?rest_route=\/wp\/v2\/posts"}],"about":[{"href":"https:\/\/savepearlharbor.com\/index.php?rest_route=\/wp\/v2\/types\/post"}],"author":[{"embeddable":true,"href":"https:\/\/savepearlharbor.com\/index.php?rest_route=\/wp\/v2\/users\/1"}],"replies":[{"embeddable":true,"href":"https:\/\/savepearlharbor.com\/index.php?rest_route=%2Fwp%2Fv2%2Fcomments&post=489110"}],"version-history":[{"count":0,"href":"https:\/\/savepearlharbor.com\/index.php?rest_route=\/wp\/v2\/posts\/489110\/revisions"}],"wp:attachment":[{"href":"https:\/\/savepearlharbor.com\/index.php?rest_route=%2Fwp%2Fv2%2Fmedia&parent=489110"}],"wp:term":[{"taxonomy":"category","embeddable":true,"href":"https:\/\/savepearlharbor.com\/index.php?rest_route=%2Fwp%2Fv2%2Fcategories&post=489110"},{"taxonomy":"post_tag","embeddable":true,"href":"https:\/\/savepearlharbor.com\/index.php?rest_route=%2Fwp%2Fv2%2Ftags&post=489110"}],"curies":[{"name":"wp","href":"https:\/\/api.w.org\/{rel}","templated":true}]}}