<?xml version="1.0" encoding="UTF-8"?>
<!DOCTYPE article PUBLIC "-//NLM//DTD JATS (Z39.96) Journal Publishing DTD v1.1 20151215//EN" "http://jats.nlm.nih.gov/publishing/1.1/JATS-journalpublishing1.dtd">
<article xmlns:xlink="http://www.w3.org/1999/xlink" xmlns:mml="http://www.w3.org/1998/Math/MathML" xml:lang="en" article-type="research-article" dtd-version="1.1">
<front>
<journal-meta>
<journal-id journal-id-type="pmc">CMC</journal-id>
<journal-id journal-id-type="nlm-ta">CMC</journal-id>
<journal-id journal-id-type="publisher-id">CMC</journal-id>
<journal-title-group>
<journal-title>Computers, Materials &#x0026; Continua</journal-title>
</journal-title-group>
<issn pub-type="epub">1546-2226</issn>
<issn pub-type="ppub">1546-2218</issn>
<publisher>
<publisher-name>Tech Science Press</publisher-name>
<publisher-loc>USA</publisher-loc>
</publisher>
</journal-meta>
<article-meta>
<article-id pub-id-type="publisher-id">82357</article-id>
<article-id pub-id-type="doi">10.32604/cmc.2026.082357</article-id>
<article-categories>
<subj-group subj-group-type="heading">
<subject>Article</subject>
</subj-group>
</article-categories>
<title-group>
<article-title>IG-Mamba: Isoline-Guided Evolutionary State Space Model for Physics-Informed Underwater Image Restoration</article-title>
<alt-title alt-title-type="left-running-head">IG-Mamba: Isoline-Guided Evolutionary State Space Model for Physics-Informed Underwater Image Restoration</alt-title>
<alt-title alt-title-type="right-running-head">IG-Mamba: Isoline-Guided Evolutionary State Space Model for Physics-Informed Underwater Image Restoration</alt-title>
</title-group>
<contrib-group>
<contrib id="author-1" contrib-type="author">
<name name-style="western"><surname>Xiang</surname><given-names>Yiqiao</given-names></name><xref ref-type="aff" rid="aff-1">1</xref></contrib>
<contrib id="author-2" contrib-type="author" corresp="yes">
<name name-style="western"><surname>Zhou</surname><given-names>Jingchun</given-names></name><xref ref-type="aff" rid="aff-1">1</xref><xref ref-type="aff" rid="aff-2">2</xref><email>zhoujingchun@dlmu.edu.cn</email></contrib>
<contrib id="author-3" contrib-type="author">
<name name-style="western"><surname>Liu</surname><given-names>Ruijie</given-names></name><xref ref-type="aff" rid="aff-1">1</xref></contrib>
<contrib id="author-4" contrib-type="author">
<name name-style="western"><surname>Zhang</surname><given-names>Dehuan</given-names></name><xref ref-type="aff" rid="aff-1">1</xref></contrib>
<aff id="aff-1"><label>1</label><institution>College of Information Science and Technology, Dalian Maritime University</institution>, <addr-line>Dalian</addr-line>, <country>China</country></aff>
<aff id="aff-2"><label>2</label><institution>State Key Laboratory of Ocean Sensing, ZJU-Hangzhou Global Scientific and Technological Innovation Center, Zhejiang University</institution>, <addr-line>Hangzhou</addr-line>, <country>China</country></aff>
</contrib-group>
<author-notes>
<corresp id="cor1"><label>&#x002A;</label>Corresponding Author: Jingchun Zhou. Email: <email>zhoujingchun@dlmu.edu.cn</email></corresp>
</author-notes>
<pub-date date-type="collection" publication-format="electronic">
<year>2026</year>
</pub-date>
<pub-date date-type="pub" publication-format="electronic">
<day>23</day><month>07</month><year>2026</year>
</pub-date>
<volume>88</volume>
<issue>3</issue>
<elocation-id>66</elocation-id>
<history>
<date date-type="received">
<day>16</day>
<month>03</month>
<year>2026</year>
</date>
<date date-type="accepted">
<day>18</day>
<month>05</month>
<year>2026</year>
</date>
</history>
<permissions>
<copyright-statement>&#x00A9; 2026 The Authors. Published by Tech Science Press.</copyright-statement>
<copyright-year>2026</copyright-year>
<copyright-holder>The Authors</copyright-holder>
<license xlink:href="https://creativecommons.org/licenses/by/4.0/">
<license-p>This work is licensed under a <ext-link ext-link-type="uri" xlink:type="simple" xlink:href="https://creativecommons.org/licenses/by/4.0/">Creative Commons Attribution 4.0 International License</ext-link>, which permits unrestricted use, distribution, and reproduction in any medium, provided the original work is properly cited.</license-p>
</license>
</permissions>
<self-uri content-type="pdf" xlink:href="TSP_CMC_82357.pdf"></self-uri>
<abstract>
<p>Underwater imagery is degraded by depth-dependent absorption and scattering, which often introduce color casts and contrast attenuation. Although recent Vision Mamba models provide efficient long-range dependency modeling, their conventional 2D scanning patterns are not explicitly designed to exploit the depth-correlated structure of underwater degradation and may therefore weaken geometry-aware feature dependencies. To address this limitation, we propose Isoline-Guided Evolutionary Mamba (IG-Mamba), a physics-inspired framework that uses a depth-correlated potential prior to organize state-space token propagation. Specifically, we introduce a Topology-Preserving Isoline Scanning mechanism. By leveraging a geometric prior, this mechanism quantizes the scene into discrete iso-potential strata to guide Mamba sequences along geometry-aware orders, thereby preserving local spatial topology while establishing depth-aware long-range dependencies. Furthermore, a Potential Field Evolution method is developed to mitigate the domain discrepancy between terrestrial geometric priors and underwater optical attenuation. Driven by the reconstruction objective, the network adaptively refines the raw geometric anchor into a restoration-oriented optical potential field via a learned residual map. Finally, features are modulated by a transmission-inspired gate motivated by the Jaffe-McGlamery transmission term, enabling spatially adaptive feature reweighting under scattering-dominant conditions. Extensive experiments demonstrate that IG-Mamba achieves strong performance across multiple benchmarks, offering a physics-grounded perspective for dependency modeling in underwater vision.</p>
</abstract>
<kwd-group kwd-group-type="author">
<kwd>Underwater image enhancement</kwd>
<kwd>state space model</kwd>
<kwd>isoline scanning</kwd>
<kwd>physics-informed learning</kwd>
</kwd-group>
<funding-group>
<award-group id="awg1">
<funding-source>National Natural Science Foundation of China</funding-source>
<award-id>62301105</award-id>
</award-group>
<award-group id="awg2">
<funding-source>State Key Laboratory of Ocean Sensing</funding-source>
<award-id>OSKF-2025M04</award-id>
</award-group>
<award-group id="awg3">
<funding-source>Natural Science Foundation-General Program</funding-source>
<award-id>2025-MSLH-113</award-id>
</award-group>
<award-group id="awg4">
<funding-source>Fundamental Research Funds for the Central Universities</funding-source>
<award-id>3132026240</award-id>
<award-id>3132025268</award-id>
</award-group>
<award-group id="awg5">
<funding-source>Liaoning Provincial Department of Education</funding-source>
<award-id>LJ212510151021</award-id>
</award-group>
</funding-group>
</article-meta>
</front>
<body>
<sec id="s1">
<label>1</label>
<title>Introduction</title>
<p>Underwater Image Enhancement (UIE) is essential for marine operations such as autonomous underwater vehicles (AUVs) and biological observation [<xref ref-type="bibr" rid="ref-1">1</xref>&#x2013;<xref ref-type="bibr" rid="ref-3">3</xref>]. According to the Jaffe-McGlamery model, underwater degradation is mainly caused by wavelength-dependent absorption and scattering, both of which are correlated with scene depth [<xref ref-type="bibr" rid="ref-4">4</xref>&#x2013;<xref ref-type="bibr" rid="ref-6">6</xref>]. Although recent data-driven UIE methods achieve strong performance, many of them still organize feature interactions primarily on the 2D image plane, without explicitly exploiting the depth-conditioned structure of underwater degradation.</p>
<p>To examine this structure, we compute a scattering-related degradation proxy on paired UIEB images:<disp-formula id="ueqn-1"><mml:math id="mml-ueqn-1" display="block"><mml:mtable columnalign="right left right left right left right left right left right left" rowspacing="3pt" columnspacing="0em 2em 0em 2em 0em 2em 0em 2em 0em 2em 0em" displaystyle="true"><mml:mtr><mml:mtd><mml:msub><mml:mrow><mml:mover><mml:mi>D</mml:mi><mml:mo stretchy="false">&#x007E;</mml:mo></mml:mover></mml:mrow><mml:mi>i</mml:mi></mml:msub><mml:mo stretchy="false">(</mml:mo><mml:mi>x</mml:mi><mml:mo stretchy="false">)</mml:mo><mml:mo>=</mml:mo><mml:mo movablelimits="true" form="prefix">max</mml:mo><mml:mspace width="negativethinmathspace" /><mml:mrow><mml:mo>(</mml:mo><mml:mi>C</mml:mi><mml:mo stretchy="false">(</mml:mo><mml:msub><mml:mi>G</mml:mi><mml:mrow><mml:mi>&#x03C3;</mml:mi></mml:mrow></mml:msub><mml:mo stretchy="false">(</mml:mo><mml:msubsup><mml:mi>Y</mml:mi><mml:mi>i</mml:mi><mml:mrow><mml:mrow><mml:mrow><mml:mtext>gt</mml:mtext></mml:mrow></mml:mrow></mml:mrow></mml:msubsup><mml:mo stretchy="false">)</mml:mo><mml:mo stretchy="false">)</mml:mo><mml:mo stretchy="false">(</mml:mo><mml:mi>x</mml:mi><mml:mo stretchy="false">)</mml:mo><mml:mo>&#x2212;</mml:mo><mml:mi>C</mml:mi><mml:mo stretchy="false">(</mml:mo><mml:msub><mml:mi>G</mml:mi><mml:mrow><mml:mi>&#x03C3;</mml:mi></mml:mrow></mml:msub><mml:mo stretchy="false">(</mml:mo><mml:msubsup><mml:mi>Y</mml:mi><mml:mi>i</mml:mi><mml:mrow><mml:mrow><mml:mtext>in</mml:mtext></mml:mrow></mml:mrow></mml:msubsup><mml:mo stretchy="false">)</mml:mo><mml:mo stretchy="false">)</mml:mo><mml:mo stretchy="false">(</mml:mo><mml:mi>x</mml:mi><mml:mo stretchy="false">)</mml:mo><mml:mo>,</mml:mo><mml:mspace width="thinmathspace" /><mml:mn>0</mml:mn><mml:mo>)</mml:mo></mml:mrow><mml:mo>,</mml:mo></mml:mtd></mml:mtr></mml:mtable></mml:math></disp-formula>where <inline-formula id="ieqn-1"><mml:math id="mml-ieqn-1"><mml:msubsup><mml:mi>Y</mml:mi><mml:mi>i</mml:mi><mml:mrow><mml:mrow><mml:mi mathvariant="normal">g</mml:mi><mml:mi mathvariant="normal">t</mml:mi></mml:mrow></mml:mrow></mml:msubsup></mml:math></inline-formula> and <inline-formula id="ieqn-2"><mml:math id="mml-ieqn-2"><mml:msubsup><mml:mi>Y</mml:mi><mml:mi>i</mml:mi><mml:mrow><mml:mrow><mml:mi mathvariant="normal">i</mml:mi><mml:mi mathvariant="normal">n</mml:mi></mml:mrow></mml:mrow></mml:msubsup></mml:math></inline-formula> denote the luminance representations of the reference and degraded images, respectively, <inline-formula id="ieqn-3"><mml:math id="mml-ieqn-3"><mml:msub><mml:mi>G</mml:mi><mml:mrow><mml:mi>&#x03C3;</mml:mi></mml:mrow></mml:msub></mml:math></inline-formula> denotes Gaussian low-pass filtering, and <inline-formula id="ieqn-4"><mml:math id="mml-ieqn-4"><mml:mi>C</mml:mi><mml:mo stretchy="false">(</mml:mo><mml:mo>&#x22C5;</mml:mo><mml:mo stretchy="false">)</mml:mo></mml:math></inline-formula> is a local-contrast operator implemented by local standard deviation. For aggregation, the proxy is min&#x2013;max normalized within each image. We then divide pixels into five near-to-far quantile bins according to an external depth-correlated prior and compute the bin-wise mean of the normalized proxy. As shown in <xref ref-type="fig" rid="fig-1">Fig. 1</xref>, the proxy generally increases from near to far, suggesting that underwater degradation contains depth-conditioned structure rather than only unstructured corruption.</p>
<fig id="fig-1">
<label>Figure 1</label>
<caption>
<title>Depth-conditioned degradation statistics on UIEB. Pixels are divided into five near-to-far bins according to the external depth-correlated prior. The increasing trend suggests that underwater degradation exhibits depth-conditioned statistics.</title>
</caption>
<graphic mimetype="image" mime-subtype="tif" xlink:href="CMC_82357-fig-1.tif"/>
</fig>
<p>Mamba [<xref ref-type="bibr" rid="ref-7">7</xref>] provides efficient long-range sequence modeling, but when applied to vision tasks, its effectiveness critically depends on how 2D tokens are organized into 1D sequences. Existing Vision Mamba variants usually adopt raster, cross, or other image-plane scan orders, which are not explicitly aligned with depth-related underwater degradation. This motivates us to move from coordinate-based scanning to geometry-aware sequence organization.</p>
<p>We propose IG-Mamba, a physics-inspired state-space framework for underwater image restoration. Its key component is Topology-Preserving Isoline Scanning, which partitions pixels into discrete iso-potential strata using a depth-correlated potential prior. Within each stratum, stable sorting preserves the original row-wise or column-wise local order, reducing local fragmentation while enabling state propagation along geometry-aware orders. Since terrestrial geometric priors may not match underwater optical degradation, we further introduce Potential Field Evolution to refine the coarse geometry anchor into a restoration-oriented optical potential, followed by transmission-inspired feature modulation for scattering-dominant regions.</p>
<p>The main contributions are summarized as follows:<list list-type="bullet">
<list-item>
<p>We propose IG-Mamba, which uses a depth-correlated potential field to organize the scanning topology of visual SSMs for underwater image restoration.</p></list-item>
<list-item>
<p>We present Topology-Preserving Isoline Scanning, which guides Mamba sequences along geometry-aware potential strata and preserves local 2D order within each stratum through stable sorting.</p></list-item>
<list-item>
<p>We introduce Potential Field Evolution and transmission-inspired feature modulation to calibrate coarse terrestrial geometry priors into restoration-oriented optical potentials.</p></list-item>
<list-item>
<p>Extensive experiments, degraded-prior stress tests, efficiency analysis, and bin-number sensitivity studies demonstrate the effectiveness and practical feasibility of IG-Mamba.</p></list-item>
</list></p>
</sec>
<sec id="s2">
<label>2</label>
<title>Related Work</title>
<sec id="s2_1">
<label>2.1</label>
<title>Underwater Image Enhancement</title>
<p>Underwater image enhancement (UIE) has been extensively studied to counteract the severe color shifts and contrast loss caused by light absorption and scattering [<xref ref-type="bibr" rid="ref-8">8</xref>]. According to the classic Jaffe-McGlamery optical model [<xref ref-type="bibr" rid="ref-9">9</xref>], underwater imaging is a depth-dependent process where the degradation is highly correlated with the distance between the camera and the reflective object. Early approaches heavily relied on hand-crafted physical priors, such as the Dark Channel Prior (DCP) [<xref ref-type="bibr" rid="ref-10">10</xref>], to estimate transmission maps. However, these prior-based methods often suffer from limited generalization due to their brittle assumptions.</p>
<p>In recent years, learning-based UIE methods have dominated the field [<xref ref-type="bibr" rid="ref-11">11</xref>&#x2013;<xref ref-type="bibr" rid="ref-14">14</xref>]. While achieving significant improvements in visual quality, many learning-based UIE methods still formulate restoration mainly as an image-to-image mapping problem. Physical priors, when used, are often introduced as auxiliary inputs, supervision signals, or output-level constraints, rather than being deeply embedded into the feature interaction process.</p>
</sec>
<sec id="s2_2">
<label>2.2</label>
<title>Vision State Space Models</title>
<p>State Space Models (SSMs) derive from continuous-time systems, and Mamba [<xref ref-type="bibr" rid="ref-7">7</xref>] excels at modeling long-term dependencies. Applying 1D sequence models to 2D vision problems generally requires a spatial scanning strategy, such as the four-directional cross-scan in SS2D [<xref ref-type="bibr" rid="ref-15">15</xref>].</p>
<p>However, existing methods generally treat Mamba as a general-purpose feature extractor and organize 2D feature maps into 1D sequences along predetermined image-plane paths. Such coordinate-defined scanning orders do not explicitly align the state-propagation path of SSMs with the depth-dependent optical degradation in underwater scenes. Consequently, scanning feature maps across different degradation states with a fixed image-plane order may be less aligned with underwater degradation statistics.</p>
</sec>
<sec id="s2_3">
<label>2.3</label>
<title>Physics-Driven Deep Learning and Depth Priors</title>
<p>Recognizing the critical role of 3D physical geometry in underwater degradation, previous UIE studies have attempted to incorporate depth or physical priors through multi-task depth estimation, feature concatenation, region-level guidance, or physical constraints [<xref ref-type="bibr" rid="ref-16">16</xref>&#x2013;<xref ref-type="bibr" rid="ref-18">18</xref>]. These methods demonstrate that geometry-related cues are useful for UIE, but the prior usually participates as an auxiliary input, an additional supervision signal, or a constraint on the restored image. In contrast, IG-Mamba uses a depth-correlated potential field to define the token ordering of the SSM sequence itself. Thus, the prior does not merely enrich the input representation; it organizes the state-propagation topology of Mamba through iso-potential routing. This distinction shifts depth guidance from feature injection to geometry-aware sequence organization. To avoid training an underwater depth estimator from scratch, our implementation uses DepthAnythingV2 [<xref ref-type="bibr" rid="ref-19">19</xref>] as a default prior provider, while later experiments further show that the framework is not tied to this specific estimator.</p>
</sec>
</sec>
<sec id="s3">
<label>3</label>
<title>Methodology</title>
<sec id="s3_1">
<label>3.1</label>
<title>Overall Architecture</title>
<p>The overall architecture of the proposed Isoline-Guided Evolutionary Mamba (IG-Mamba) is illustrated in <xref ref-type="fig" rid="fig-2">Fig. 2</xref>. Unlike many UIE networks that rely solely on degraded color images, IG-Mamba adopts a physics-informed dual-input strategy. Given a degraded underwater image <inline-formula id="ieqn-5"><mml:math id="mml-ieqn-5"><mml:msub><mml:mi>I</mml:mi><mml:mrow><mml:mtext>in</mml:mtext></mml:mrow></mml:msub><mml:mo>&#x2208;</mml:mo><mml:msup><mml:mrow><mml:mi mathvariant="double-struck">R</mml:mi></mml:mrow><mml:mrow><mml:mi>B</mml:mi><mml:mo>&#x00D7;</mml:mo><mml:mn>3</mml:mn><mml:mo>&#x00D7;</mml:mo><mml:mi>H</mml:mi><mml:mo>&#x00D7;</mml:mo><mml:mi>W</mml:mi></mml:mrow></mml:msup></mml:math></inline-formula> and its corresponding depth-correlated geometry prior <inline-formula id="ieqn-6"><mml:math id="mml-ieqn-6"><mml:msub><mml:mi>P</mml:mi><mml:mrow><mml:mtext>in</mml:mtext></mml:mrow></mml:msub><mml:mo>&#x2208;</mml:mo><mml:msup><mml:mrow><mml:mi mathvariant="double-struck">R</mml:mi></mml:mrow><mml:mrow><mml:mi>B</mml:mi><mml:mo>&#x00D7;</mml:mo><mml:mn>1</mml:mn><mml:mo>&#x00D7;</mml:mo><mml:mi>H</mml:mi><mml:mo>&#x00D7;</mml:mo><mml:mi>W</mml:mi></mml:mrow></mml:msup></mml:math></inline-formula>, our objective is to reconstruct the enhanced image <inline-formula id="ieqn-7"><mml:math id="mml-ieqn-7"><mml:msub><mml:mi>I</mml:mi><mml:mrow><mml:mtext>out</mml:mtext></mml:mrow></mml:msub></mml:math></inline-formula>. Specifically, <inline-formula id="ieqn-8"><mml:math id="mml-ieqn-8"><mml:msub><mml:mi>I</mml:mi><mml:mrow><mml:mtext>in</mml:mtext></mml:mrow></mml:msub></mml:math></inline-formula> is processed by a convolutional stem to extract initial visual features <inline-formula id="ieqn-9"><mml:math id="mml-ieqn-9"><mml:msub><mml:mi>X</mml:mi><mml:mn>0</mml:mn></mml:msub><mml:mo>&#x2208;</mml:mo><mml:msup><mml:mrow><mml:mi mathvariant="double-struck">R</mml:mi></mml:mrow><mml:mrow><mml:mi>B</mml:mi><mml:mo>&#x00D7;</mml:mo><mml:mi>C</mml:mi><mml:mo>&#x00D7;</mml:mo><mml:mi>H</mml:mi><mml:mo>&#x00D7;</mml:mo><mml:mi>W</mml:mi></mml:mrow></mml:msup></mml:math></inline-formula>. Concurrently, <inline-formula id="ieqn-10"><mml:math id="mml-ieqn-10"><mml:msub><mml:mi>P</mml:mi><mml:mrow><mml:mtext>in</mml:mtext></mml:mrow></mml:msub></mml:math></inline-formula> serves as the initial potential field <inline-formula id="ieqn-11"><mml:math id="mml-ieqn-11"><mml:msub><mml:mi>P</mml:mi><mml:mn>0</mml:mn></mml:msub></mml:math></inline-formula> for the routing and modulation stream. As shown in <xref ref-type="fig" rid="fig-2">Fig. 2</xref>, the backbone follows a multi-stage encoder-decoder structure: the encoder consists of <monospace>IGStage</monospace> modules, while the decoder uses corresponding <monospace>IGStage_up</monospace> modules with skip connections from the encoder. Each stage contains stacked <monospace>IGBlock</monospace> units for geometry-guided restoration, followed by spatial downsampling or upsampling operations. For notation consistency, <monospace>IGBlock</monospace> denotes the basic unit containing IsolineSort-guided Mamba modeling, local detail preservation, potential updating, and transmission-inspired modulation.</p>
<fig id="fig-2">
<label>Figure 2</label>
<caption>
<title>Overall architecture of IG-Mamba. An external geometry prior initializes the potential stream, which is propagated together with the visual feature stream through the encoder-decoder backbone. <monospace>IGStage</monospace> and <monospace>IGStage_up</monospace> denote the encoder and decoder stages. (<bold>a</bold>) <monospace>IGBlock</monospace> combines local convolution, horizontal and vertical IsolineSort-guided Mamba branches, gated local-detail preservation, <italic>P</italic>-updating, and transmission-inspired feature modulation. (<bold>b</bold>) The <italic>P</italic>-updater predicts <inline-formula id="ieqn-12"><mml:math id="mml-ieqn-12"><mml:mi mathvariant="normal">&#x0394;</mml:mi><mml:mi>P</mml:mi></mml:math></inline-formula> and updates the potential by <inline-formula id="ieqn-13"><mml:math id="mml-ieqn-13"><mml:msub><mml:mi>P</mml:mi><mml:mrow><mml:mrow><mml:mi mathvariant="normal">n</mml:mi><mml:mi mathvariant="normal">e</mml:mi><mml:mi mathvariant="normal">w</mml:mi></mml:mrow></mml:mrow></mml:msub><mml:mo>=</mml:mo><mml:mi>P</mml:mi><mml:mo>+</mml:mo><mml:mi mathvariant="normal">&#x0394;</mml:mi><mml:mi>P</mml:mi></mml:math></inline-formula>. (<bold>c</bold>) IsolineSort quantizes <italic>P</italic> into <inline-formula id="ieqn-14"><mml:math id="mml-ieqn-14"><mml:msub><mml:mi>P</mml:mi><mml:mi>q</mml:mi></mml:msub></mml:math></inline-formula> and uses stable argsort to build geometry-aware token sequences while preserving local order within each iso-potential stratum.</title>
</caption>
<graphic mimetype="image" mime-subtype="tif" xlink:href="CMC_82357-fig-2.tif"/>
</fig>
<p>For clarity, <inline-formula id="ieqn-15"><mml:math id="mml-ieqn-15"><mml:msub><mml:mi>P</mml:mi><mml:mn>0</mml:mn></mml:msub></mml:math></inline-formula>, <italic>P</italic>, <inline-formula id="ieqn-16"><mml:math id="mml-ieqn-16"><mml:msub><mml:mi>P</mml:mi><mml:mi>q</mml:mi></mml:msub></mml:math></inline-formula>, and <inline-formula id="ieqn-17"><mml:math id="mml-ieqn-17"><mml:msub><mml:mi>P</mml:mi><mml:mrow><mml:mtext>new</mml:mtext></mml:mrow></mml:msub></mml:math></inline-formula> denote different computational states of the same potential stream. <inline-formula id="ieqn-18"><mml:math id="mml-ieqn-18"><mml:msub><mml:mi>P</mml:mi><mml:mn>0</mml:mn></mml:msub></mml:math></inline-formula> is the external initial geometry prior, <italic>P</italic> is the current potential field received by a block, <inline-formula id="ieqn-19"><mml:math id="mml-ieqn-19"><mml:msub><mml:mi>P</mml:mi><mml:mi>q</mml:mi></mml:msub></mml:math></inline-formula> is the quantized potential used only for isoline sorting, and <inline-formula id="ieqn-20"><mml:math id="mml-ieqn-20"><mml:msub><mml:mi>P</mml:mi><mml:mrow><mml:mtext>new</mml:mtext></mml:mrow></mml:msub><mml:mo>=</mml:mo><mml:mi>P</mml:mi><mml:mo>+</mml:mo><mml:mi mathvariant="normal">&#x0394;</mml:mi><mml:mi>P</mml:mi></mml:math></inline-formula> is the evolved potential used for subsequent propagation and physical modulation. The framework does not assume <inline-formula id="ieqn-21"><mml:math id="mml-ieqn-21"><mml:msub><mml:mi>P</mml:mi><mml:mn>0</mml:mn></mml:msub></mml:math></inline-formula> to be metrically accurate depth; it serves as a coarse spatial geometry anchor.</p>
<p>To maintain a stable scale space for the potential field across hierarchical levels, the potential stream employs average pooling for spatial downsampling and bilinear interpolation for upsampling. During skip connections, element-wise mean fusion is applied (<inline-formula id="ieqn-22"><mml:math id="mml-ieqn-22"><mml:msub><mml:mi>P</mml:mi><mml:mrow><mml:mtext>out</mml:mtext></mml:mrow></mml:msub><mml:mo>=</mml:mo><mml:mfrac><mml:mrow><mml:msub><mml:mi>P</mml:mi><mml:mrow><mml:mtext>up</mml:mtext></mml:mrow></mml:msub><mml:mo>+</mml:mo><mml:msub><mml:mi>P</mml:mi><mml:mrow><mml:mtext>skip</mml:mtext></mml:mrow></mml:msub></mml:mrow><mml:mn>2</mml:mn></mml:mfrac></mml:math></inline-formula>) instead of standard addition to prevent value inflation beyond a stable numerical range.</p>
</sec>
<sec id="s3_2">
<label>3.2</label>
<title>Physics-Guided Isoline Routing Mamba</title>
<p><bold>Optical Triangle and Radiative-Transfer-Motivated State Propagation.</bold> In underwater environments, light propagation follows the optical triangle principle: light travels from the source to the object (distance <inline-formula id="ieqn-23"><mml:math id="mml-ieqn-23"><mml:msub><mml:mi>d</mml:mi><mml:mn>1</mml:mn></mml:msub></mml:math></inline-formula>), reflects, and returns to the camera (distance <inline-formula id="ieqn-24"><mml:math id="mml-ieqn-24"><mml:msub><mml:mi>d</mml:mi><mml:mn>2</mml:mn></mml:msub></mml:math></inline-formula>). Based on the Jaffe-McGlamery model [<xref ref-type="bibr" rid="ref-9">9</xref>], visual deterioration in a stationary state is primarily governed by the camera-to-object distance. Recently, most Vision Mamba models have converted two-dimensional spatial information into one-dimensional sequences using fixed scanning pathways, such as rasterization. However, spatial-agnostic scanning may connect pixels with different depth-related degradation states in the SSM sequence, thereby weakening local structural integrity. To resolve this issue, we propose the <italic>Topology-Preserving Isoline Scanning</italic> scheme based on the one-dimensional line-of-sight Radiative Transfer Equation (RTE) theory. In underwater optics, the propagation of light at a distance <inline-formula id="ieqn-25"><mml:math id="mml-ieqn-25"><mml:mi>z</mml:mi></mml:math></inline-formula> obeys the following law:<disp-formula id="eqn-1"><label>(1)</label><mml:math id="mml-eqn-1" display="block"><mml:mtable columnalign="right left right left right left right left right left right left" rowspacing="3pt" columnspacing="0em 2em 0em 2em 0em 2em 0em 2em 0em 2em 0em" displaystyle="true"><mml:mtr><mml:mtd /><mml:mtd><mml:mfrac><mml:mrow><mml:mi>d</mml:mi><mml:mi>L</mml:mi></mml:mrow><mml:mrow><mml:mi>d</mml:mi><mml:mi>z</mml:mi></mml:mrow></mml:mfrac><mml:mo>=</mml:mo><mml:mo>&#x2212;</mml:mo><mml:mi>c</mml:mi><mml:mo>&#x22C5;</mml:mo><mml:mi>L</mml:mi><mml:mo stretchy="false">(</mml:mo><mml:mi>z</mml:mi><mml:mo stretchy="false">)</mml:mo><mml:mo>+</mml:mo><mml:mi>S</mml:mi><mml:mo stretchy="false">(</mml:mo><mml:mi>z</mml:mi><mml:mo stretchy="false">)</mml:mo><mml:mo>,</mml:mo></mml:mtd></mml:mtr></mml:mtable></mml:math></disp-formula></p>
<p>Here, <inline-formula id="ieqn-26"><mml:math id="mml-ieqn-26"><mml:mi>z</mml:mi></mml:math></inline-formula> denotes the physical optical-path coordinate. Notably, the continuous-time SSM equation has a similar recursive form:<disp-formula id="eqn-2"><label>(2)</label><mml:math id="mml-eqn-2" display="block"><mml:mtable columnalign="right left right left right left right left right left right left" rowspacing="3pt" columnspacing="0em 2em 0em 2em 0em 2em 0em 2em 0em 2em 0em" displaystyle="true"><mml:mtr><mml:mtd /><mml:mtd><mml:msup><mml:mi>h</mml:mi><mml:mi mathvariant="normal">&#x2032;</mml:mi></mml:msup><mml:mo stretchy="false">(</mml:mo><mml:mi>t</mml:mi><mml:mo stretchy="false">)</mml:mo><mml:mo>=</mml:mo><mml:mi>A</mml:mi><mml:mi>h</mml:mi><mml:mo stretchy="false">(</mml:mo><mml:mi>t</mml:mi><mml:mo stretchy="false">)</mml:mo><mml:mo>+</mml:mo><mml:mi>B</mml:mi><mml:mi>x</mml:mi><mml:mo stretchy="false">(</mml:mo><mml:mi>t</mml:mi><mml:mo stretchy="false">)</mml:mo><mml:mo>.</mml:mo></mml:mtd></mml:mtr></mml:mtable></mml:math></disp-formula></p>
<p>In IG-Mamba, the learned potential <italic>P</italic> is not treated as the physical optical-path coordinate. Instead, it is used as a depth-correlated ordering variable to group pixels into near-to-far iso-potential strata. By grouping pixels into near-to-far iso-potential strata, IsolineSort constructs geometry-aware token sequences in which later hidden states can aggregate information from preceding tokens with related optical paths or degradation states. This provides a physics-inspired inductive bias analogous to radiative transfer, where radiance accumulates along an optical path through attenuation and scattering.</p>
<p><bold>Orthogonal Isoline Routing.</bold> To approximate this ordered accumulation over discrete tensors while better preserving local two-dimensional structure, we quantize the potential field <italic>P</italic> into discrete ordinal iso-potential indices with a bin granularity controlled by <italic>K</italic>. Specifically, <inline-formula id="ieqn-27"><mml:math id="mml-ieqn-27"><mml:mtext>Quantize</mml:mtext><mml:mo stretchy="false">(</mml:mo><mml:mo>&#x22C5;</mml:mo><mml:mo>,</mml:mo><mml:mi>K</mml:mi><mml:mo stretchy="false">)</mml:mo></mml:math></inline-formula> maps potential values into discrete ordinal indices, where pixels assigned to the same index are treated as belonging to the same iso-potential stratum. To address potential directional bias, the input features are decomposed by channel into horizontal (<inline-formula id="ieqn-28"><mml:math id="mml-ieqn-28"><mml:msub><mml:mi>X</mml:mi><mml:mi>h</mml:mi></mml:msub></mml:math></inline-formula>) and vertical (<inline-formula id="ieqn-29"><mml:math id="mml-ieqn-29"><mml:msub><mml:mi>X</mml:mi><mml:mi>v</mml:mi></mml:msub></mml:math></inline-formula>) branches, respectively.</p>
<p>The resulting operation, termed <italic>IsolineSort</italic>, applies stable sorting according to the quantized potential rather than exact metric depth. Stability ensures that pixels within the same iso-potential stratum retain their original row-wise and column-wise local orders in the <inline-formula id="ieqn-30"><mml:math id="mml-ieqn-30"><mml:msub><mml:mi>X</mml:mi><mml:mi>h</mml:mi></mml:msub></mml:math></inline-formula> and <inline-formula id="ieqn-31"><mml:math id="mml-ieqn-31"><mml:msub><mml:mi>X</mml:mi><mml:mi>v</mml:mi></mml:msub></mml:math></inline-formula> branches, respectively. These visual tokens are then flattened to form one-dimensional, isoline-ordered sequences. After independent state-space evolution by <inline-formula id="ieqn-32"><mml:math id="mml-ieqn-32"><mml:msub><mml:mtext>Mamba</mml:mtext><mml:mi>H</mml:mi></mml:msub></mml:math></inline-formula> and <inline-formula id="ieqn-33"><mml:math id="mml-ieqn-33"><mml:msub><mml:mtext>Mamba</mml:mtext><mml:mi>V</mml:mi></mml:msub></mml:math></inline-formula>, the sequences are mapped back to their original 2D spatial grids by inverse sorting and then concatenated. The detailed dual-route algorithm is shown in Algorithm 1. Here, <inline-formula id="ieqn-34"><mml:math id="mml-ieqn-34"><mml:mo stretchy="false">(</mml:mo><mml:mo>&#x22C5;</mml:mo><mml:msup><mml:mo stretchy="false">)</mml:mo><mml:mrow><mml:mrow><mml:mi mathvariant="sans-serif">T</mml:mi></mml:mrow></mml:mrow></mml:msup></mml:math></inline-formula> denotes transposition of the two spatial dimensions <italic>H</italic> and <italic>W</italic>.</p>
<fig id="fig-7">
<graphic mimetype="image" mime-subtype="tif" xlink:href="CMC_82357-fig-7.tif"/>
</fig>
</sec>
<sec id="s3_3">
<label>3.3</label>
<title>Parallel Local Detail Branch</title>
<p>While the isoline-guided SSM captures macro-level optical dependencies, it might inherently overlook localized, high-frequency textures that are orthogonal to the depth gradient (e.g., coral textures). We introduce a parallel Local Detail Branch utilizing a <inline-formula id="ieqn-48"><mml:math id="mml-ieqn-48"><mml:mn>3</mml:mn><mml:mo>&#x00D7;</mml:mo><mml:mn>3</mml:mn></mml:math></inline-formula> depth-wise convolution and Channel Attention:<disp-formula id="eqn-3"><label>(3)</label><mml:math id="mml-eqn-3" display="block"><mml:mtable columnalign="right left right left right left right left right left right left" rowspacing="3pt" columnspacing="0em 2em 0em 2em 0em 2em 0em 2em 0em 2em 0em" displaystyle="true"><mml:mtr><mml:mtd /><mml:mtd><mml:msub><mml:mi>X</mml:mi><mml:mrow><mml:mi>l</mml:mi><mml:mi>o</mml:mi><mml:mi>c</mml:mi><mml:mi>a</mml:mi><mml:mi>l</mml:mi></mml:mrow></mml:msub><mml:mo>=</mml:mo><mml:msub><mml:mrow><mml:mi>&#x0210B;</mml:mi></mml:mrow><mml:mrow><mml:mi>c</mml:mi><mml:mi>o</mml:mi><mml:mi>n</mml:mi><mml:mi>v</mml:mi></mml:mrow></mml:msub><mml:mo stretchy="false">(</mml:mo><mml:mi>X</mml:mi><mml:mo stretchy="false">)</mml:mo><mml:mo>&#x2299;</mml:mo><mml:mi>&#x03C3;</mml:mi><mml:mo stretchy="false">(</mml:mo><mml:msub><mml:mrow><mml:mi>&#x0210B;</mml:mi></mml:mrow><mml:mrow><mml:mi>g</mml:mi><mml:mi>a</mml:mi><mml:mi>t</mml:mi><mml:mi>e</mml:mi></mml:mrow></mml:msub><mml:mo stretchy="false">(</mml:mo><mml:mrow><mml:mtext>GAP</mml:mtext></mml:mrow><mml:mo stretchy="false">(</mml:mo><mml:msub><mml:mrow><mml:mi>&#x0210B;</mml:mi></mml:mrow><mml:mrow><mml:mi>c</mml:mi><mml:mi>o</mml:mi><mml:mi>n</mml:mi><mml:mi>v</mml:mi></mml:mrow></mml:msub><mml:mo stretchy="false">(</mml:mo><mml:mi>X</mml:mi><mml:mo stretchy="false">)</mml:mo><mml:mo stretchy="false">)</mml:mo><mml:mo stretchy="false">)</mml:mo><mml:mo stretchy="false">)</mml:mo><mml:mo>,</mml:mo></mml:mtd></mml:mtr></mml:mtable></mml:math></disp-formula></p>
<p>The aggregated feature is fused via a learnable scaling parameter <inline-formula id="ieqn-49"><mml:math id="mml-ieqn-49"><mml:mi>&#x03B1;</mml:mi></mml:math></inline-formula>: <inline-formula id="ieqn-50"><mml:math id="mml-ieqn-50"><mml:msub><mml:mi>X</mml:mi><mml:mrow><mml:mi>f</mml:mi><mml:mi>u</mml:mi><mml:mi>s</mml:mi><mml:mi>e</mml:mi><mml:mi>d</mml:mi></mml:mrow></mml:msub><mml:mo>=</mml:mo><mml:mtext>FFN</mml:mtext><mml:mo stretchy="false">(</mml:mo><mml:msub><mml:mi>Y</mml:mi><mml:mrow><mml:mi>r</mml:mi><mml:mi>o</mml:mi><mml:mi>u</mml:mi><mml:mi>t</mml:mi><mml:mi>e</mml:mi></mml:mrow></mml:msub><mml:mo>+</mml:mo><mml:msub><mml:mi>X</mml:mi><mml:mrow><mml:mi>l</mml:mi><mml:mi>o</mml:mi><mml:mi>c</mml:mi><mml:mi>a</mml:mi><mml:mi>l</mml:mi></mml:mrow></mml:msub><mml:mo>&#x22C5;</mml:mo><mml:mi>&#x03B1;</mml:mi><mml:mo stretchy="false">)</mml:mo><mml:mo>+</mml:mo><mml:mi>X</mml:mi></mml:math></inline-formula>. This dual-branch design supports both global geometry-aware modeling and local visual fidelity.</p>
</sec>
<sec id="s3_4">
<label>3.4</label>
<title>Potential Field Evolution and Transmission-Inspired Modulation</title>
<p>Although pre-trained foundation models can provide useful geometric priors, a non-negligible domain gap remains between in-air geometric depth and underwater optical degradation, which is affected by turbidity, backscattering, and wavelength-dependent absorption. The same depth-dependent degradation principle also motivates a complementary modulation path: while geometry-aware routing uses the potential to define the state-propagation order, the Jaffe-McGlamery transmission term <inline-formula id="ieqn-51"><mml:math id="mml-ieqn-51"><mml:mi>t</mml:mi><mml:mo stretchy="false">(</mml:mo><mml:mi>x</mml:mi><mml:mo stretchy="false">)</mml:mo><mml:mo>=</mml:mo><mml:msup><mml:mi>e</mml:mi><mml:mrow><mml:mo>&#x2212;</mml:mo><mml:mi>&#x03B2;</mml:mi><mml:mi>d</mml:mi><mml:mo stretchy="false">(</mml:mo><mml:mi>x</mml:mi><mml:mo stretchy="false">)</mml:mo></mml:mrow></mml:msup></mml:math></inline-formula> suggests that depth-correlated geometry can also provide an attenuation-related signal for feature reweighting.</p>
<p>To implement this idea while reducing the geometry-optics domain gap, IG-Mamba proposes the <italic>Potential Field Evolution (<inline-formula id="ieqn-52"><mml:math id="mml-ieqn-52"><mml:mi mathvariant="normal">&#x0394;</mml:mi><mml:mi>P</mml:mi></mml:math></inline-formula>)</italic> mechanism (Algorithm 2). During the forward pass, the coarse geometric prior is refined by a learned residual correction <inline-formula id="ieqn-53"><mml:math id="mml-ieqn-53"><mml:mi mathvariant="normal">&#x0394;</mml:mi><mml:mi>P</mml:mi></mml:math></inline-formula> to obtain the updated optical potential field <inline-formula id="ieqn-54"><mml:math id="mml-ieqn-54"><mml:msub><mml:mi>P</mml:mi><mml:mrow><mml:mrow><mml:mi mathvariant="normal">n</mml:mi><mml:mi mathvariant="normal">e</mml:mi><mml:mi mathvariant="normal">w</mml:mi></mml:mrow></mml:mrow></mml:msub><mml:mo>&#x2208;</mml:mo><mml:msup><mml:mrow><mml:mi mathvariant="double-struck">R</mml:mi></mml:mrow><mml:mrow><mml:mi>B</mml:mi><mml:mo>&#x00D7;</mml:mo><mml:mn>1</mml:mn><mml:mo>&#x00D7;</mml:mo><mml:mi>H</mml:mi><mml:mo>&#x00D7;</mml:mo><mml:mi>W</mml:mi></mml:mrow></mml:msup></mml:math></inline-formula>. Here, <inline-formula id="ieqn-55"><mml:math id="mml-ieqn-55"><mml:msub><mml:mi>P</mml:mi><mml:mrow><mml:mrow><mml:mi mathvariant="normal">n</mml:mi><mml:mi mathvariant="normal">e</mml:mi><mml:mi mathvariant="normal">w</mml:mi></mml:mrow></mml:mrow></mml:msub></mml:math></inline-formula> should be interpreted as a restoration-oriented, degradation-aware optical potential rather than a metrically accurate depth map. A lightweight convolutional mapping then expands <inline-formula id="ieqn-56"><mml:math id="mml-ieqn-56"><mml:msub><mml:mi>P</mml:mi><mml:mrow><mml:mrow><mml:mi mathvariant="normal">n</mml:mi><mml:mi mathvariant="normal">e</mml:mi><mml:mi mathvariant="normal">w</mml:mi></mml:mrow></mml:mrow></mml:msub></mml:math></inline-formula> into a spatial-channel transmission-inspired gate <inline-formula id="ieqn-57"><mml:math id="mml-ieqn-57"><mml:msub><mml:mi>W</mml:mi><mml:mrow><mml:mrow><mml:mi mathvariant="normal">m</mml:mi><mml:mi mathvariant="normal">o</mml:mi><mml:mi mathvariant="normal">d</mml:mi></mml:mrow></mml:mrow></mml:msub><mml:mo>&#x2208;</mml:mo><mml:mo stretchy="false">[</mml:mo><mml:mn>0</mml:mn><mml:mo>,</mml:mo><mml:mn>1</mml:mn><mml:msup><mml:mo stretchy="false">]</mml:mo><mml:mrow><mml:mi>B</mml:mi><mml:mo>&#x00D7;</mml:mo><mml:mi>C</mml:mi><mml:mo>&#x00D7;</mml:mo><mml:mi>H</mml:mi><mml:mo>&#x00D7;</mml:mo><mml:mi>W</mml:mi></mml:mrow></mml:msup></mml:math></inline-formula> for feature reweighting.</p>
<p>The aggregated visual features are then modulated through element-wise multiplication, <inline-formula id="ieqn-58"><mml:math id="mml-ieqn-58"><mml:msub><mml:mi>X</mml:mi><mml:mrow><mml:mrow><mml:mi mathvariant="normal">f</mml:mi><mml:mi mathvariant="normal">u</mml:mi><mml:mi mathvariant="normal">s</mml:mi><mml:mi mathvariant="normal">e</mml:mi><mml:mi mathvariant="normal">d</mml:mi></mml:mrow></mml:mrow></mml:msub><mml:mo>&#x2299;</mml:mo><mml:msub><mml:mi>W</mml:mi><mml:mrow><mml:mrow><mml:mi mathvariant="normal">m</mml:mi><mml:mi mathvariant="normal">o</mml:mi><mml:mi mathvariant="normal">d</mml:mi></mml:mrow></mml:mrow></mml:msub></mml:math></inline-formula>. This differentiable branch also plays an optimization role beyond its physical motivation. Since the <monospace>argsort</monospace> operations in our isoline routing are non-differentiable and prevent direct gradient propagation through sorting indices, <inline-formula id="ieqn-59"><mml:math id="mml-ieqn-59"><mml:msub><mml:mi>W</mml:mi><mml:mrow><mml:mrow><mml:mi mathvariant="normal">m</mml:mi><mml:mi mathvariant="normal">o</mml:mi><mml:mi mathvariant="normal">d</mml:mi></mml:mrow></mml:mrow></mml:msub></mml:math></inline-formula> provides an additional differentiable optimization path. Consequently, improper modulation, such as over-suppressing a bright foreground region, increases the reconstruction loss, which backpropagates through <inline-formula id="ieqn-60"><mml:math id="mml-ieqn-60"><mml:msub><mml:mi>W</mml:mi><mml:mrow><mml:mrow><mml:mi mathvariant="normal">m</mml:mi><mml:mi mathvariant="normal">o</mml:mi><mml:mi mathvariant="normal">d</mml:mi></mml:mrow></mml:mrow></mml:msub></mml:math></inline-formula> to adaptively calibrate the potential field. Radiometric compensation is mainly handled by the main restoration stream and reconstruction head, while the modulation branch provides transmission-inspired reweighting for scattering-dominant feature responses and encourages the potential stream to retain its association with depth-correlated geometry.</p>
<p>Finally, this design forms a restoration-driven correction loop: the routing path uses the current potential to organize state propagation, while the modulation path allows the reconstruction objective to adapt the potential through a continuous feature-reweighting branch. This encourages <inline-formula id="ieqn-61"><mml:math id="mml-ieqn-61"><mml:mi mathvariant="normal">&#x0394;</mml:mi><mml:mi>P</mml:mi></mml:math></inline-formula> to adjust the raw geometric prior toward a restoration-oriented optical potential without requiring auxiliary depth supervision.</p>
<fig id="fig-8">
<graphic mimetype="image" mime-subtype="tif" xlink:href="CMC_82357-fig-8.tif"/>
</fig>
</sec>
<sec id="s3_5">
<label>3.5</label>
<title>Supervision Strategy of Potential Field</title>
<p>We train IG-Mamba end-to-end with a reconstruction-driven potential evolution strategy. The external prior <inline-formula id="ieqn-72"><mml:math id="mml-ieqn-72"><mml:msub><mml:mi>P</mml:mi><mml:mn>0</mml:mn></mml:msub></mml:math></inline-formula> generated by DepthAnythingV2 [<xref ref-type="bibr" rid="ref-19">19</xref>] is used only to initialize the potential stream as a coarse geometry anchor. Through IsolineSort-guided routing and the differentiable transmission-inspired modulation branch, the image restoration objective optimizes the potential stream and encourages <inline-formula id="ieqn-73"><mml:math id="mml-ieqn-73"><mml:mi mathvariant="normal">&#x0394;</mml:mi><mml:mi>P</mml:mi></mml:math></inline-formula> to adapt the raw geometric prior into a restoration-oriented optical potential that is more consistent with underwater optical degradation.</p>
<p>The overall loss is entirely specified as the <inline-formula id="ieqn-74"><mml:math id="mml-ieqn-74"><mml:msub><mml:mi>L</mml:mi><mml:mn>1</mml:mn></mml:msub></mml:math></inline-formula> reconstruction loss of the image:<disp-formula id="eqn-4"><label>(4)</label><mml:math id="mml-eqn-4" display="block"><mml:msub><mml:mrow><mml:mi>&#x02112;</mml:mi></mml:mrow><mml:mrow><mml:mrow><mml:mtext>total</mml:mtext></mml:mrow></mml:mrow></mml:msub><mml:mo>=</mml:mo><mml:msub><mml:mrow><mml:mi>&#x02112;</mml:mi></mml:mrow><mml:mrow><mml:mrow><mml:mtext>1</mml:mtext></mml:mrow></mml:mrow></mml:msub><mml:mo>=</mml:mo><mml:mfrac><mml:mn>1</mml:mn><mml:mi>N</mml:mi></mml:mfrac><mml:mo fence="false" stretchy="false">&#x2016;</mml:mo><mml:msub><mml:mi>I</mml:mi><mml:mrow><mml:mi>p</mml:mi><mml:mi>r</mml:mi><mml:mi>e</mml:mi><mml:mi>d</mml:mi></mml:mrow></mml:msub><mml:mo>&#x2212;</mml:mo><mml:msub><mml:mi>I</mml:mi><mml:mrow><mml:mi>g</mml:mi><mml:mi>t</mml:mi></mml:mrow></mml:msub><mml:msub><mml:mo fence="false" stretchy="false">&#x2016;</mml:mo><mml:mn>1</mml:mn></mml:msub></mml:math></disp-formula>where <italic>N</italic> is the total number of pixels in <inline-formula id="ieqn-75"><mml:math id="mml-ieqn-75"><mml:msub><mml:mi>I</mml:mi><mml:mrow><mml:mi>p</mml:mi><mml:mi>r</mml:mi><mml:mi>e</mml:mi><mml:mi>d</mml:mi></mml:mrow></mml:msub></mml:math></inline-formula>.</p>
<p>This depth-supervision-free design allows the potential field to be optimized for image restoration rather than being forced to fit potentially inaccurate metric depth labels.</p>
</sec>
</sec>
<sec id="s4">
<label>4</label>
<title>Experiments</title>
<sec id="s4_1">
<label>4.1</label>
<title>Experimental Setup</title>
<p><bold>Datasets.</bold> Training and in-domain validation are conducted on UIEB [<xref ref-type="bibr" rid="ref-11">11</xref>]: 800 paired images are used for training (UIEB-T800), 90 paired images for full-reference validation (UIEB-V90), and 60 challenging images without references for UIEB-C60 evaluation. Cross-domain non-reference evaluation is further conducted on U45 [<xref ref-type="bibr" rid="ref-20">20</xref>], UCCS [<xref ref-type="bibr" rid="ref-21">21</xref>], MABLs [<xref ref-type="bibr" rid="ref-22">22</xref>], UFO120 [<xref ref-type="bibr" rid="ref-23">23</xref>], and Z700 [<xref ref-type="bibr" rid="ref-24">24</xref>], covering different color casts, turbidity levels, and real underwater scenes.</p>
<p><bold>Implementation Details.</bold> IG-Mamba is implemented in PyTorch and trained end-to-end on an NVIDIA vGPU (32 GB). We use Adam with <inline-formula id="ieqn-76"><mml:math id="mml-ieqn-76"><mml:msub><mml:mi>&#x03B2;</mml:mi><mml:mn>1</mml:mn></mml:msub><mml:mo>=</mml:mo><mml:mn>0.9</mml:mn></mml:math></inline-formula>, <inline-formula id="ieqn-77"><mml:math id="mml-ieqn-77"><mml:msub><mml:mi>&#x03B2;</mml:mi><mml:mn>2</mml:mn></mml:msub><mml:mo>=</mml:mo><mml:mn>0.999</mml:mn></mml:math></inline-formula>, an initial learning rate of <inline-formula id="ieqn-78"><mml:math id="mml-ieqn-78"><mml:mn>1</mml:mn><mml:mo>&#x00D7;</mml:mo><mml:msup><mml:mn>10</mml:mn><mml:mrow><mml:mo>&#x2212;</mml:mo><mml:mn>4</mml:mn></mml:mrow></mml:msup></mml:math></inline-formula>, and cosine annealing to <inline-formula id="ieqn-79"><mml:math id="mml-ieqn-79"><mml:mn>1</mml:mn><mml:mo>&#x00D7;</mml:mo><mml:msup><mml:mn>10</mml:mn><mml:mrow><mml:mo>&#x2212;</mml:mo><mml:mn>6</mml:mn></mml:mrow></mml:msup></mml:math></inline-formula>. The network is trained for 600 epochs with a batch size of 4. Images are randomly cropped to <inline-formula id="ieqn-80"><mml:math id="mml-ieqn-80"><mml:mn>256</mml:mn><mml:mo>&#x00D7;</mml:mo><mml:mn>256</mml:mn></mml:math></inline-formula> and augmented by random horizontal/vertical flips. The external prior is provided as an input guidance field, and the model is optimized by image reconstruction loss.</p>
<p>The default external prior is generated by DepthAnythingV2 and precomputed before restoration experiments. It is normalized as a depth-correlated potential input and is not used as a supervised depth label. In the degraded-prior robustness study, we keep the trained IG-Mamba checkpoint fixed and only replace or corrupt <inline-formula id="ieqn-81"><mml:math id="mml-ieqn-81"><mml:msub><mml:mi>P</mml:mi><mml:mn>0</mml:mn></mml:msub></mml:math></inline-formula> at inference time, so the results evaluate test-time tolerance to imperfect prior providers rather than robustness obtained from prior augmentation.</p>
<p>For competing methods, we use official results or released models when available; otherwise, reproduced outputs are evaluated with the same metric scripts and testing splits. This keeps the full-reference and non-reference comparisons under a consistent protocol.</p>
<p><bold>Evaluation Metrics.</bold> PSNR [<xref ref-type="bibr" rid="ref-25">25</xref>] and SSIM [<xref ref-type="bibr" rid="ref-26">26</xref>] are used for full-reference evaluation, and RMSE is additionally reported on UIEB-V90. UCIQE [<xref ref-type="bibr" rid="ref-27">27</xref>] and URanker [<xref ref-type="bibr" rid="ref-28">28</xref>] are used for non-reference evaluation. Higher values indicate better performance for PSNR, SSIM, UCIQE, and URanker, while lower values indicate better performance for RMSE.</p>
</sec>
<sec id="s4_2">
<label>4.2</label>
<title>Comparison with State-of-the-Art Methods</title>
<p>We compare IG-Mamba with classic prior-/CNN-based methods (WaterNet [<xref ref-type="bibr" rid="ref-11">11</xref>], UWCNN [<xref ref-type="bibr" rid="ref-12">12</xref>], Ucolor [<xref ref-type="bibr" rid="ref-13">13</xref>], SPDF [<xref ref-type="bibr" rid="ref-29">29</xref>]), Transformer/advanced restoration methods (U-shape [<xref ref-type="bibr" rid="ref-14">14</xref>], AST [<xref ref-type="bibr" rid="ref-30">30</xref>], VQCNIR [<xref ref-type="bibr" rid="ref-31">31</xref>]), and recent competitive models including FMambaIR [<xref ref-type="bibr" rid="ref-32">32</xref>], UDNet [<xref ref-type="bibr" rid="ref-33">33</xref>], and CDF-UIE [<xref ref-type="bibr" rid="ref-34">34</xref>].</p>
<p><bold>Quantitative Evaluation.</bold> <xref ref-type="table" rid="table-1">Tables 1</xref> and <xref ref-type="table" rid="table-2">2</xref> report the quantitative comparison results. Benefiting from geometry-aware routing and global state-space modeling, IG-Mamba achieves consistently strong performance across the evaluated datasets. Compared with SSM-based methods that rely on generic image-plane scanning paths, such as FMambaIR [<xref ref-type="bibr" rid="ref-32">32</xref>], IG-Mamba obtains higher PSNR on UIEB-V90, suggesting that organizing tokens by a depth-correlated potential field provides a useful inductive bias for depth-dependent underwater degradation.</p>
<table-wrap id="table-1">
<label>Table 1</label>
<caption>
<title>Quantitative comparison of different methods on UIEB-V90 and UIEB-C60 datasets. The 1st, 2nd, and 3rd best results are highlighted in <styled-content style-type="color" style="background-color: #FFD9D9;"><bold>red</bold></styled-content>, <styled-content style-type="color" style="background-color: #D9D9FF;"><underline>blue</underline></styled-content>, and <styled-content style-type="color" style="background-color: #D9FFD9;">green</styled-content> backgrounds, respectively. Ties are assigned the same rank.</title>
</caption>
<table>
<colgroup>
<col align="center"/>
<col align="center"/>
<col align="center"/>
<col align="center"/>
<col align="center"/>
<col align="center"/>
<col align="center"/>
<col align="center"/>
<col align="center"/> </colgroup>
<thead>
<tr>
<th rowspan="2">Method</th>
<th rowspan="2">Venue</th>
<th colspan="5">UIEB-V90</th>
<th colspan="2">UIEB-C60</th>
</tr>
<tr>
<th>PSNR<inline-formula id="ieqn-83"><mml:math id="mml-ieqn-83"><mml:mo stretchy="false">&#x2191;</mml:mo></mml:math></inline-formula></th>
<th>SSIM<inline-formula id="ieqn-84"><mml:math id="mml-ieqn-84"><mml:mo stretchy="false">&#x2191;</mml:mo></mml:math></inline-formula></th>
<th>RMSE<inline-formula id="ieqn-85"><mml:math id="mml-ieqn-85"><mml:mo stretchy="false">&#x2193;</mml:mo></mml:math></inline-formula></th>
<th>URanker<inline-formula id="ieqn-86"><mml:math id="mml-ieqn-86"><mml:mo stretchy="false">&#x2191;</mml:mo></mml:math></inline-formula></th>
<th>UCIQE<inline-formula id="ieqn-87"><mml:math id="mml-ieqn-87"><mml:mo stretchy="false">&#x2191;</mml:mo></mml:math></inline-formula></th>
<th>URanker<inline-formula id="ieqn-88"><mml:math id="mml-ieqn-88"><mml:mo stretchy="false">&#x2191;</mml:mo></mml:math></inline-formula></th>
<th>UCIQE<inline-formula id="ieqn-89"><mml:math id="mml-ieqn-89"><mml:mo stretchy="false">&#x2191;</mml:mo></mml:math></inline-formula></th>
</tr>
</thead>
<tbody>
<tr>
<td>WaterNet [<xref ref-type="bibr" rid="ref-11">11</xref>]</td>
<td>TIP&#x2019;19</td>
<td>17.349</td>
<td>0.813</td>
<td>0.144</td>
<td>1.471</td>
<td>0.606</td>
<td>0.931</td>
<td><styled-content style-type="color" style="background-color: #FFD9D9;"><bold>0.591</bold></styled-content></td>
</tr>
<tr>
<td>UWCNN [<xref ref-type="bibr" rid="ref-12">12</xref>]</td>
<td>PR&#x2019;20</td>
<td>17.985</td>
<td>0.844</td>
<td>0.134</td>
<td>1.309</td>
<td>0.554</td>
<td>0.499</td>
<td>0.530</td>
</tr>
<tr>
<td>Ucolor [<xref ref-type="bibr" rid="ref-13">13</xref>]</td>
<td>TIP&#x2019;21</td>
<td>20.962</td>
<td>0.864</td>
<td>0.097</td>
<td>1.534</td>
<td>0.591</td>
<td>0.841</td>
<td>0.565</td>
</tr>
<tr>
<td>SPDF [<xref ref-type="bibr" rid="ref-29">29</xref>]</td>
<td>TCSVT&#x2019;23</td>
<td>20.331</td>
<td>0.828</td>
<td>0.103</td>
<td>1.433</td>
<td>0.570</td>
<td>1.034</td>
<td>0.569</td>
</tr>
<tr>
<td>U-shape [<xref ref-type="bibr" rid="ref-14">14</xref>]</td>
<td>TIP&#x2019;23</td>
<td>20.462</td>
<td>0.792</td>
<td>0.100</td>
<td>1.665</td>
<td>0.576</td>
<td>1.059</td>
<td>0.553</td>
</tr>
<tr>
<td>VQCNIR [<xref ref-type="bibr" rid="ref-31">31</xref>]</td>
<td>AAAI&#x2019;24</td>
<td><styled-content style-type="color" style="background-color: #D9D9FF;"><underline>23.330</underline></styled-content></td>
<td><styled-content style-type="color" style="background-color: #FFD9D9;"><bold>0.914</bold></styled-content></td>
<td><styled-content style-type="color" style="background-color: #D9D9FF;"><underline>0.074</underline></styled-content></td>
<td><styled-content style-type="color" style="background-color: #D9FFD9;">2.069</styled-content></td>
<td><styled-content style-type="color" style="background-color: #D9D9FF;"><underline>0.615</underline></styled-content></td>
<td><styled-content style-type="color" style="background-color: #D9FFD9;">1.446</styled-content></td>
<td><styled-content style-type="color" style="background-color: #D9D9FF;"><underline>0.588</underline></styled-content></td>
</tr>
<tr>
<td>AST [<xref ref-type="bibr" rid="ref-30">30</xref>]</td>
<td>CVPR&#x2019;24</td>
<td>21.695</td>
<td><styled-content style-type="color" style="background-color: #D9FFD9;">0.869</styled-content></td>
<td>0.089</td>
<td>1.791</td>
<td>0.586</td>
<td>1.042</td>
<td>0.558</td>
</tr>
<tr>
<td>FMambaIR [<xref ref-type="bibr" rid="ref-32">32</xref>]</td>
<td>TGRS&#x2019;25</td>
<td><styled-content style-type="color" style="background-color: #D9FFD9;">22.305</styled-content></td>
<td>0.865</td>
<td><styled-content style-type="color" style="background-color: #D9FFD9;">0.083</styled-content></td>
<td><styled-content style-type="color" style="background-color: #D9D9FF;"><underline>2.096</underline></styled-content></td>
<td><styled-content style-type="color" style="background-color: #D9FFD9;">0.613</styled-content></td>
<td><styled-content style-type="color" style="background-color: #D9D9FF;"><underline>1.489</underline></styled-content></td>
<td><styled-content style-type="color" style="background-color: #D9D9FF;"><underline>0.588</underline></styled-content></td>
</tr>
<tr>
<td>UDNet [<xref ref-type="bibr" rid="ref-33">33</xref>]</td>
<td>ESWA&#x2019;25</td>
<td>19.017</td>
<td>0.768</td>
<td>0.118</td>
<td>1.216</td>
<td>0.585</td>
<td>0.618</td>
<td>0.568</td>
</tr>
<tr>
<td>CDF-UIE [<xref ref-type="bibr" rid="ref-34">34</xref>]</td>
<td>TGRS&#x2019;25</td>
<td>21.699</td>
<td>0.808</td>
<td>0.088</td>
<td>2.042</td>
<td>0.603</td>
<td>1.387</td>
<td><styled-content style-type="color" style="background-color: #D9FFD9;">0.576</styled-content></td>
</tr>
<tr>
<td><bold>IG-Mamba</bold></td>
<td>Ours</td>
<td><styled-content style-type="color" style="background-color: #FFD9D9;"><bold>24.521</bold></styled-content></td>
<td><styled-content style-type="color" style="background-color: #D9D9FF;"><underline>0.907</underline></styled-content></td>
<td><styled-content style-type="color" style="background-color: #FFD9D9;"><bold>0.069</bold></styled-content></td>
<td><styled-content style-type="color" style="background-color: #FFD9D9;"><bold>2.210</bold></styled-content></td>
<td><styled-content style-type="color" style="background-color: #FFD9D9;"><bold>0.621</bold></styled-content></td>
<td><styled-content style-type="color" style="background-color: #FFD9D9;"><bold>1.573</bold></styled-content></td>
<td><styled-content style-type="color" style="background-color: #FFD9D9;"><bold>0.591</bold></styled-content></td>
</tr>
</tbody>
</table>
</table-wrap><table-wrap id="table-2">
<label>Table 2</label>
<caption>
<title>Quantitative comparison of different methods on U45, Z700, MABLs, UFO120, and UCCS datasets. The 1st, 2nd, and 3rd best results are highlighted in <styled-content style-type="color" style="background-color: #FFD9D9;"><bold>red</bold></styled-content>, <styled-content style-type="color" style="background-color: #D9D9FF;"><underline>blue</underline></styled-content>, and <styled-content style-type="color" style="background-color: #D9FFD9;">green</styled-content> backgrounds, respectively. Ties are assigned the same rank.</title>
</caption>
<table>
<colgroup>
<col align="center"/>
<col align="center"/>
<col align="center"/>
<col align="center"/>
<col align="center"/>
<col align="center"/>
<col align="center"/>
<col align="center"/>
<col align="center"/>
<col align="center"/>
<col align="center"/>
<col align="center"/> </colgroup>
<thead>
<tr>
<th align="center" rowspan="2">Method</th>
<th align="center" rowspan="2">Venue</th>
<th align="center" colspan="2">U45</th>
<th align="center" colspan="2">Z700</th>
<th align="center" colspan="2">MABLs</th>
<th align="center" colspan="2">UFO120</th>
<th align="center" colspan="2">UCCS</th>
</tr>
<tr>
<th>URanker<inline-formula id="ieqn-90"><mml:math id="mml-ieqn-90"><mml:mo stretchy="false">&#x2191;</mml:mo></mml:math></inline-formula></th>
<th>UCIQE<inline-formula id="ieqn-91"><mml:math id="mml-ieqn-91"><mml:mo stretchy="false">&#x2191;</mml:mo></mml:math></inline-formula></th>
<th>URanker<inline-formula id="ieqn-92"><mml:math id="mml-ieqn-92"><mml:mo stretchy="false">&#x2191;</mml:mo></mml:math></inline-formula></th>
<th>UCIQE<inline-formula id="ieqn-93"><mml:math id="mml-ieqn-93"><mml:mo stretchy="false">&#x2191;</mml:mo></mml:math></inline-formula></th>
<th>URanker<inline-formula id="ieqn-94"><mml:math id="mml-ieqn-94"><mml:mo stretchy="false">&#x2191;</mml:mo></mml:math></inline-formula></th>
<th>UCIQE<inline-formula id="ieqn-95"><mml:math id="mml-ieqn-95"><mml:mo stretchy="false">&#x2191;</mml:mo></mml:math></inline-formula></th>
<th>URanker<inline-formula id="ieqn-96"><mml:math id="mml-ieqn-96"><mml:mo stretchy="false">&#x2191;</mml:mo></mml:math></inline-formula></th>
<th>UCIQE<inline-formula id="ieqn-97"><mml:math id="mml-ieqn-97"><mml:mo stretchy="false">&#x2191;</mml:mo></mml:math></inline-formula></th>
<th>URanker<inline-formula id="ieqn-98"><mml:math id="mml-ieqn-98"><mml:mo stretchy="false">&#x2191;</mml:mo></mml:math></inline-formula></th>
<th>UCIQE<inline-formula id="ieqn-99"><mml:math id="mml-ieqn-99"><mml:mo stretchy="false">&#x2191;</mml:mo></mml:math></inline-formula></th>
</tr>
</thead>
<tbody>
<tr>
<td>WaterNet [<xref ref-type="bibr" rid="ref-11">11</xref>]</td>
<td>TIP&#x2019;19</td>
<td>1.330</td>
<td>0.594</td>
<td>1.707</td>
<td>0.559</td>
<td>1.448</td>
<td><styled-content style-type="color" style="background-color: #FFD9D9;"><bold>0.600</bold></styled-content></td>
<td>1.748</td>
<td>0.615</td>
<td>1.216</td>
<td><styled-content style-type="color" style="background-color: #FFD9D9;"><bold>0.565</bold></styled-content></td>
</tr>
<tr>
<td>UWCNN [<xref ref-type="bibr" rid="ref-12">12</xref>]</td>
<td>PR&#x2019;20</td>
<td>1.513</td>
<td>0.554</td>
<td>1.495</td>
<td>0.503</td>
<td>1.030</td>
<td>0.532</td>
<td>1.893</td>
<td>0.579</td>
<td>0.983</td>
<td>0.501</td>
</tr>
<tr>
<td>Ucolor [<xref ref-type="bibr" rid="ref-13">13</xref>]</td>
<td>TIP&#x2019;21</td>
<td>1.471</td>
<td>0.586</td>
<td>1.895</td>
<td>0.543</td>
<td>1.255</td>
<td>0.573</td>
<td>1.754</td>
<td>0.612</td>
<td>1.090</td>
<td>0.553</td>
</tr>
<tr>
<td>SPDF [<xref ref-type="bibr" rid="ref-29">29</xref>]</td>
<td>TCSVT&#x2019;23</td>
<td>1.284</td>
<td>0.565</td>
<td>1.492</td>
<td>0.527</td>
<td>1.413</td>
<td>0.560</td>
<td>1.530</td>
<td>0.572</td>
<td>1.212</td>
<td>0.508</td>
</tr>
<tr>
<td>U-shape [<xref ref-type="bibr" rid="ref-14">14</xref>]</td>
<td>TIP&#x2019;23</td>
<td>1.502</td>
<td>0.572</td>
<td>1.976</td>
<td>0.553</td>
<td>1.514</td>
<td>0.561</td>
<td>1.678</td>
<td>0.592</td>
<td>1.178</td>
<td>0.528</td>
</tr>
<tr>
<td>VQCNIR [<xref ref-type="bibr" rid="ref-31">31</xref>]</td>
<td>AAAI&#x2019;24</td>
<td><styled-content style-type="color" style="background-color: #FFD9D9;"><bold>2.067</bold></styled-content></td>
<td><styled-content style-type="color" style="background-color: #D9D9FF;"><underline>0.616</underline></styled-content></td>
<td><styled-content style-type="color" style="background-color: #D9D9FF;"><underline>2.106</underline></styled-content></td>
<td><styled-content style-type="color" style="background-color: #D9D9FF;"><underline>0.577</underline></styled-content></td>
<td><styled-content style-type="color" style="background-color: #D9D9FF;"><underline>1.812</underline></styled-content></td>
<td><styled-content style-type="color" style="background-color: #D9FFD9;">0.592</styled-content></td>
<td><styled-content style-type="color" style="background-color: #D9FFD9;">2.122</styled-content></td>
<td><styled-content style-type="color" style="background-color: #D9D9FF;"><underline>0.635</underline></styled-content></td>
<td><styled-content style-type="color" style="background-color: #D9D9FF;"><underline>1.371</underline></styled-content></td>
<td>0.550</td>
</tr>
<tr>
<td>AST [<xref ref-type="bibr" rid="ref-30">30</xref>]</td>
<td>CVPR&#x2019;24</td>
<td>1.776</td>
<td>0.589</td>
<td>1.807</td>
<td>0.544</td>
<td>1.363</td>
<td>0.564</td>
<td>1.938</td>
<td>0.608</td>
<td>0.827</td>
<td>0.523</td>
</tr>
<tr>
<td>FMambaIR [<xref ref-type="bibr" rid="ref-32">32</xref>]</td>
<td>TGRS&#x2019;25</td>
<td>1.853</td>
<td><styled-content style-type="color" style="background-color: #D9FFD9;">0.609</styled-content></td>
<td>1.937</td>
<td><styled-content style-type="color" style="background-color: #D9FFD9;">0.574</styled-content></td>
<td><styled-content style-type="color" style="background-color: #D9FFD9;">1.811</styled-content></td>
<td>0.590</td>
<td><styled-content style-type="color" style="background-color: #FFD9D9;"><bold>2.175</bold></styled-content></td>
<td><styled-content style-type="color" style="background-color: #FFD9D9;"><bold>0.636</bold></styled-content></td>
<td>1.267</td>
<td>0.548</td>
</tr>
<tr>
<td>UDNet [<xref ref-type="bibr" rid="ref-33">33</xref>]</td>
<td>ESWA&#x2019;25</td>
<td>0.858</td>
<td>0.584</td>
<td>1.299</td>
<td>0.558</td>
<td>0.997</td>
<td>0.569</td>
<td>1.506</td>
<td>0.602</td>
<td>0.765</td>
<td><styled-content style-type="color" style="background-color: #D9D9FF;"><underline>0.555</underline></styled-content></td>
</tr>
<tr>
<td>CDF-UIE [<xref ref-type="bibr" rid="ref-34">34</xref>]</td>
<td>TGRS&#x2019;25</td>
<td><styled-content style-type="color" style="background-color: #D9FFD9;">1.915</styled-content></td>
<td>0.603</td>
<td><styled-content style-type="color" style="background-color: #D9FFD9;">2.091</styled-content></td>
<td>0.563</td>
<td>1.798</td>
<td>0.583</td>
<td><styled-content style-type="color" style="background-color: #D9D9FF;"><underline>2.134</underline></styled-content></td>
<td><styled-content style-type="color" style="background-color: #D9FFD9;">0.624</styled-content></td>
<td><styled-content style-type="color" style="background-color: #D9FFD9;">1.306</styled-content></td>
<td>0.547</td>
</tr>
<tr>
<td><bold>IG-Mamba</bold></td>
<td>Ours</td>
<td><styled-content style-type="color" style="background-color: #D9D9FF;"><underline>2.029</underline></styled-content></td>
<td><styled-content style-type="color" style="background-color: #FFD9D9;"><bold>0.618</bold></styled-content></td>
<td><styled-content style-type="color" style="background-color: #FFD9D9;"><bold>2.292</bold></styled-content></td>
<td><styled-content style-type="color" style="background-color: #FFD9D9;"><bold>0.593</bold></styled-content></td>
<td><styled-content style-type="color" style="background-color: #FFD9D9;"><bold>1.935</bold></styled-content></td>
<td><styled-content style-type="color" style="background-color: #D9D9FF;"><underline>0.599</underline></styled-content></td>
<td><styled-content style-type="color" style="background-color: #D9D9FF;"><underline>2.134</underline></styled-content></td>
<td><styled-content style-type="color" style="background-color: #D9D9FF;"><underline>0.635</underline></styled-content></td>
<td><styled-content style-type="color" style="background-color: #FFD9D9;"><bold>1.380</bold></styled-content></td>
<td><styled-content style-type="color" style="background-color: #D9FFD9;">0.554</styled-content></td>
</tr>
</tbody>
</table>
</table-wrap>
<p>In addition to strong restoration performance, the core enhancement network of IG-Mamba maintains an efficient structure, with <bold>5.32M parameters</bold> and <bold>13.02G FLOPs</bold> at a resolution of <inline-formula id="ieqn-82"><mml:math id="mml-ieqn-82"><mml:mn>256</mml:mn><mml:mo>&#x00D7;</mml:mo><mml:mn>256</mml:mn></mml:math></inline-formula>. The default external prior in our experiments is precomputed, and the framework is not tied to one specific depth estimator. In practice, the coarse geometric anchor may be provided by a lightweight learned proxy, low-resolution prior, or on-board AUV sensing sources such as stereo cameras and sonars. This deployment flexibility is further analyzed in <xref ref-type="sec" rid="s4_4">Section 4.4</xref>.</p>
 
<p><bold>Qualitative Evaluation.</bold> <xref ref-type="fig" rid="fig-3">Figs. 3</xref>&#x2013;<xref ref-type="fig" rid="fig-5">5</xref> provide visual comparisons on UIEB-V90 and cross-domain datasets. IG-Mamba reduces color casts and haze-like scattering while preserving foreground structures and object details, showing robust restoration under diverse underwater conditions.</p>
<fig id="fig-3">
<label>Figure 3</label>
<caption>
<title>Qualitative comparison on reference-based and non-reference underwater image enhancement benchmarks, including UIEB-V90, UIEB-C60, U45, Z700, MABLs, UFO120, and UCCS. Columns (<bold>a</bold>&#x2013;<bold>l</bold>) correspond to the input, WaterNet, UWCNN, Ucolor, SPDF, U-shape, VQCNIR, AST, FMambaIR, UDNet, CDF-UIE, and our IG-Mamba, respectively. IG-Mamba produces visually balanced restoration results across different underwater scenes and degradation types.</title>
</caption>
<graphic mimetype="image" mime-subtype="tif" xlink:href="CMC_82357-fig-3.tif"/>
</fig><fig id="fig-4">
<label>Figure 4</label>
<caption>
<title>Visual and color-distribution comparison on a representative underwater image. Each method is shown together with its RGB probability-density curves and color distribution. Columns (<bold>a</bold>&#x2013;<bold>l</bold>) correspond to the input, WaterNet, UWCNN, Ucolor, SPDF, U-shape, VQCNIR, AST, FMambaIR, UDNet, CDF-UIE, and our IG-Mamba, respectively. Compared with existing methods, IG-Mamba better aligns the restored color distribution with the reference while reducing underwater color casts.</title>
</caption>
<graphic mimetype="image" mime-subtype="tif" xlink:href="CMC_82357-fig-4a.tif"/>
<graphic mimetype="image" mime-subtype="tif" xlink:href="CMC_82357-fig-4b.tif"/>
</fig><fig id="fig-5">
<label>Figure 5</label>
<caption>
<title>Qualitative comparison with local magnified regions on challenging non-reference underwater images. Columns (<bold>a</bold>&#x2013;<bold>l</bold>) correspond to the input, WaterNet, UWCNN, Ucolor, SPDF, U-shape, VQCNIR, AST, FMambaIR, UDNet, CDF-UIE, and our IG-Mamba, respectively. IG-Mamba preserves clearer local structures and improves visibility under challenging underwater degradation.</title>
</caption>
<graphic mimetype="image" mime-subtype="tif" xlink:href="CMC_82357-fig-5.tif"/>
</fig>
</sec>
<sec id="s4_3">
<label>4.3</label>
<title>Robustness to Degraded External Geometry Priors</title>
<p>To address the concern that IG-Mamba may rely on an accurate external depth estimator, we conduct a controlled robustness study on UIEB-V90 by deliberately degrading the input geometry prior <inline-formula id="ieqn-100"><mml:math id="mml-ieqn-100"><mml:msub><mml:mi>P</mml:mi><mml:mn>0</mml:mn></mml:msub></mml:math></inline-formula>. Downsampled and blurred priors preserve coarse but spatially aligned geometry; noisy and masked priors simulate local estimation errors or missing regions; a constant prior represents the no-geometry fallback; shuffled and random priors represent spatially inconsistent or fully corrupted guidance. <xref ref-type="table" rid="table-3">Table 3</xref> reports the robustness results under these degraded-prior settings.</p>
<table-wrap id="table-3">
<label>Table 3</label>
<caption>
<title>Robustness of IG-Mamba to degraded external geometry priors on UIEB-V90. Each cell reports PSNR/SSIM. <inline-formula id="ieqn-101"><mml:math id="mml-ieqn-101"><mml:mi mathvariant="normal">&#x0394;</mml:mi></mml:math></inline-formula>PSNR is computed relative to the clean-prior setting. Aligned perturbations preserve the spatial layout of the prior, while non-geometric priors destroy or remove the geometry guidance.</title>
</caption>
<table>
<colgroup>
<col align="center"/>
<col align="center"/>
<col align="center"/>
<col align="center"/> </colgroup>
<thead>
<tr>
<th>Prior Setting</th>
<th>Prior Property</th>
<th>IG-Mamba</th>
<th><inline-formula id="ieqn-102"><mml:math id="mml-ieqn-102"><mml:mi mathvariant="normal">&#x0394;</mml:mi></mml:math></inline-formula>PSNR</th>
</tr>
</thead>
<tbody>
<tr>
<td>Clean prior</td>
<td>Original geometry prior</td>
<td>24.52/0.9066</td>
<td>0.00</td>
</tr>
<tr>
<td>Downsampled prior</td>
<td>Coarse but aligned</td>
<td>24.53/0.9064</td>
<td>&#x002B;0.01</td>
</tr>
<tr>
<td>Blurred prior</td>
<td>Smooth but aligned</td>
<td>24.52/0.9065</td>
<td>&#x2212;0.01</td>
</tr>
<tr>
<td>Noisy prior</td>
<td>Noisy but aligned</td>
<td>24.49/0.9058</td>
<td>&#x2212;0.03</td>
</tr>
<tr>
<td>Masked prior</td>
<td>Partially missing but aligned</td>
<td>24.51/0.9058</td>
<td>&#x2212;0.01</td>
</tr>
<tr>
<td>Constant map</td>
<td>No geometry information</td>
<td>24.14/0.9024</td>
<td>&#x2212;0.38</td>
</tr>
<tr>
<td>Shuffled map</td>
<td>Spatially misaligned geometry</td>
<td>24.17/0.9011</td>
<td>&#x2212;0.36</td>
</tr>
<tr>
<td>Random map</td>
<td>Fully corrupted guidance</td>
<td>24.08/0.8996</td>
<td>&#x2212;0.44</td>
</tr>
</tbody>
</table>
</table-wrap>
<p>The results show that IG-Mamba does not require metrically accurate depth. Moderately degraded but spatially aligned priors, such as downsampled, blurred, noisy, and masked priors, do not cause performance collapse. This suggests that a coarse relative geometry anchor is sufficient for stable iso-potential grouping. In particular, coarse or smoothed priors can remain effective because they preserve the near-to-far spatial layout while reducing high-frequency artifacts in monocular depth estimates. When the spatial layout of the prior is removed or disrupted, as in constant, shuffled, and random priors, the performance drops more noticeably, suggesting that IG-Mamba relies on spatially meaningful geometry anchors rather than metrically precise depth. Together with the subsequent Static Prior ablation study, these results indicate that potential evolution helps calibrate imperfect geometry priors into restoration-oriented optical potentials.</p>
</sec>
<sec id="s4_4">
<label>4.4</label>
<title>Efficiency and Sensitivity Analysis</title>
<p><xref ref-type="table" rid="table-4">Table 4</xref> reports the efficiency and prior-provider comparison. To address practical deployment concerns, we further evaluate the computational cost of different prior providers at <inline-formula id="ieqn-103"><mml:math id="mml-ieqn-103"><mml:mn>256</mml:mn><mml:mo>&#x00D7;</mml:mo><mml:mn>256</mml:mn></mml:math></inline-formula> resolution. The external clean prior achieves the best restoration quality, but its online prior-generation cost is much higher when the DepthAnythingV2 inference time is included. In contrast, the lightweight learned prior introduces only 0.017M parameters and 0.193G FLOPs, reducing the total latency from 180.06 to 35.26 ms while maintaining competitive restoration quality. This result indicates that IG-Mamba is not tied to a heavy monocular depth foundation model; a fast task-oriented prior proxy can also provide an effective coarse geometry anchor for deployment.</p>
<table-wrap id="table-4">
<label>Table 4</label>
<caption>
<title>Efficiency and prior-provider analysis on UIEB-V90 at <inline-formula id="ieqn-104"><mml:math id="mml-ieqn-104"><mml:mn>256</mml:mn><mml:mo>&#x00D7;</mml:mo><mml:mn>256</mml:mn></mml:math></inline-formula> resolution. Runtime is measured with batch size 1 after warm-up. For external priors, the measured DepthAnythingV2 prior-generation time is included. The restoration core is shared by all settings and contains 5.32M parameters and 13.02G FLOPs.</title>
</caption>
<table>
<colgroup>
<col align="center"/>
<col align="center"/>
<col align="center"/>
<col align="center"/>
<col align="center"/>
<col align="center"/> </colgroup>
<thead>
<tr>
<th>Prior Provider</th>
<th>Prior Params</th>
<th>Prior FLOPs</th>
<th>Prior Time</th>
<th>Total Time/FPS</th>
<th>PSNR/SSIM</th>
</tr>
</thead>
<tbody>
<tr>
<td>External clean prior</td>
<td>335.316M</td>
<td>154.807G</td>
<td>145.54 ms</td>
<td>180.06 ms/5.55</td>
<td>24.521/0.9066</td>
</tr>
<tr>
<td>Lightweight learned prior</td>
<td>0.017M</td>
<td>0.193G</td>
<td>2.13 ms</td>
<td>35.26 ms/28.36</td>
<td>24.316/0.9036</td>
</tr>
<tr>
<td>Constant fallback</td>
<td>&#x2013;</td>
<td>&#x2013;</td>
<td>0.00 ms</td>
<td>33.32 ms/30.01</td>
<td>24.143/0.9024</td>
</tr>
</tbody>
</table>
</table-wrap>
<p>We also study the sensitivity to the number of iso-potential bins <italic>K</italic>, as reported in <xref ref-type="table" rid="table-5">Table 5</xref>. The sensitivity analysis shows that the no-stratification reference (<inline-formula id="ieqn-105"><mml:math id="mml-ieqn-105"><mml:mi>K</mml:mi><mml:mo>=</mml:mo><mml:mn>1</mml:mn></mml:math></inline-formula>) is clearly inferior to valid iso-potential grouping, confirming the necessity of geometry-aware sequence organization. The best performance is obtained at <inline-formula id="ieqn-106"><mml:math id="mml-ieqn-106"><mml:mi>K</mml:mi><mml:mo>=</mml:mo><mml:mn>4</mml:mn></mml:math></inline-formula>, suggesting that underwater restoration benefits from coarse but effective iso-potential strata. When <italic>K</italic> becomes larger, PSNR decreases, likely because overly fine quantization fragments spatially coherent degradation regions and amplifies high-frequency artifacts in monocular priors. This result is consistent with the degraded-prior robustness analysis: IG-Mamba requires spatially meaningful geometry guidance, but does not rely on metrically precise depth.</p>
<table-wrap id="table-5">
<label>Table 5</label>
<caption>
<title>Sensitivity analysis of the number of iso-potential bins <italic>K</italic> on UIEB-V90. The <inline-formula id="ieqn-107"><mml:math id="mml-ieqn-107"><mml:mi>K</mml:mi><mml:mo>=</mml:mo><mml:mn>1</mml:mn></mml:math></inline-formula> row corresponds to the no-stratification reference from the ablation study, while the other settings are trained with the corresponding iso-potential bin count. Bold indicates the best result.</title>
</caption>
<table>
<colgroup>
<col align="center"/>
<col align="center"/>
<col align="center"/>
<col align="center"/>
<col align="center"/> </colgroup>
<thead>
<tr>
<th><italic>K</italic></th>
<th>1</th>
<th>4</th>
<th>16</th>
<th>64</th>
</tr>
</thead>
<tbody>
<tr>
<td>PSNR<inline-formula id="ieqn-108"><mml:math id="mml-ieqn-108"><mml:mo stretchy="false">&#x2191;</mml:mo></mml:math></inline-formula></td>
<td>23.552</td>
<td><bold>24.521</bold></td>
<td>24.206</td>
<td>24.164</td>
</tr>
<tr>
<td>SSIM<inline-formula id="ieqn-109"><mml:math id="mml-ieqn-109"><mml:mo stretchy="false">&#x2191;</mml:mo></mml:math></inline-formula></td>
<td>0.902</td>
<td><bold>0.907</bold></td>
<td>0.903</td>
<td>0.905</td>
</tr>
</tbody>
</table>
</table-wrap>
</sec>
<sec id="s4_5">
<label>4.5</label>
<title>Ablation Study</title>
<p>We conduct ablation studies on UIEB to validate the main components of IG-Mamba. <xref ref-type="table" rid="table-6">Table 6</xref> analyzes routing, stable sorting, potential evolution, and modulation through a series of variants from the depth-free baseline to the full model. Here, <bold>Model A (Baseline)</bold> disables all geometry-related components and achieves 23.552 dB PSNR, providing the reference for evaluating the remaining variants.</p>
<table-wrap id="table-6">
<label>Table 6</label>
<caption>
<title>Ablation study on UIEB. Evo. denotes potential field evolution (<inline-formula id="ieqn-110"><mml:math id="mml-ieqn-110"><mml:mi mathvariant="normal">&#x0394;</mml:mi><mml:mi>P</mml:mi></mml:math></inline-formula>), and Mod. denotes transmission-inspired feature modulation. Models A&#x2013;F progressively examine the depth-free baseline, continuous sorting, unstable sorting, routing-only design, static-prior modulation, and full IG-Mamba. Bold indicates the best result.</title>
</caption>
<table>
<colgroup>
<col align="center"/>
<col align="center"/>
<col align="center"/>
<col align="center"/>
<col align="center"/>
<col align="center"/>
<col align="center"/>
<col align="center"/> </colgroup>
<thead>
<tr>
<th>Model Variant</th>
<th>Geometry Prior</th>
<th>Isoline Bins</th>
<th>Stable Sort</th>
<th><inline-formula id="ieqn-111"><mml:math id="mml-ieqn-111"><mml:mi mathvariant="normal">&#x0394;</mml:mi><mml:mi>P</mml:mi></mml:math></inline-formula> Evo.</th>
<th>Trans. Mod.</th>
<th>PSNR<inline-formula id="ieqn-112"><mml:math id="mml-ieqn-112"><mml:mo stretchy="false">&#x2191;</mml:mo></mml:math></inline-formula></th>
<th>SSIM<inline-formula id="ieqn-113"><mml:math id="mml-ieqn-113"><mml:mo stretchy="false">&#x2191;</mml:mo></mml:math></inline-formula></th>
</tr>
</thead>
<tbody>
<tr>
<td>Model A (Baseline)</td>
<td><inline-formula id="ieqn-114"><mml:math id="mml-ieqn-114"><mml:mo>&#x00D7;</mml:mo></mml:math></inline-formula></td>
<td><inline-formula id="ieqn-115"><mml:math id="mml-ieqn-115"><mml:mo>&#x00D7;</mml:mo></mml:math></inline-formula></td>
<td><inline-formula id="ieqn-116"><mml:math id="mml-ieqn-116"><mml:mo>&#x00D7;</mml:mo></mml:math></inline-formula></td>
<td><inline-formula id="ieqn-117"><mml:math id="mml-ieqn-117"><mml:mo>&#x00D7;</mml:mo></mml:math></inline-formula></td>
<td><inline-formula id="ieqn-118"><mml:math id="mml-ieqn-118"><mml:mo>&#x00D7;</mml:mo></mml:math></inline-formula></td>
<td>23.552</td>
<td>0.902</td>
</tr>
<tr>
<td>Model B (Continuous Potential)</td>
<td>&#x2713;</td>
<td><inline-formula id="ieqn-119"><mml:math id="mml-ieqn-119"><mml:mo>&#x00D7;</mml:mo></mml:math></inline-formula></td>
<td>&#x2713;</td>
<td>&#x2713;</td>
<td>&#x2713;</td>
<td>23.957</td>
<td>0.898</td>
</tr>
<tr>
<td>Model C (Unstable Sort)</td>
<td>&#x2713;</td>
<td>&#x2713;</td>
<td><inline-formula id="ieqn-120"><mml:math id="mml-ieqn-120"><mml:mo>&#x00D7;</mml:mo></mml:math></inline-formula></td>
<td>&#x2713;</td>
<td>&#x2713;</td>
<td>23.991</td>
<td>0.904</td>
</tr>
<tr>
<td>Model D (Routing Only)</td>
<td>&#x2713;</td>
<td>&#x2713;</td>
<td>&#x2713;</td>
<td><inline-formula id="ieqn-121"><mml:math id="mml-ieqn-121"><mml:mo>&#x00D7;</mml:mo></mml:math></inline-formula></td>
<td><inline-formula id="ieqn-122"><mml:math id="mml-ieqn-122"><mml:mo>&#x00D7;</mml:mo></mml:math></inline-formula></td>
<td>23.899</td>
<td>0.905</td>
</tr>
<tr>
<td>Model E (Static Prior)</td>
<td>&#x2713;</td>
<td>&#x2713;</td>
<td>&#x2713;</td>
<td><inline-formula id="ieqn-123"><mml:math id="mml-ieqn-123"><mml:mo>&#x00D7;</mml:mo></mml:math></inline-formula></td>
<td>&#x2713;</td>
<td>23.682</td>
<td>0.903</td>
</tr>
<tr>
<td><bold>Model F (IG-Mamba)</bold></td>
<td>&#x2713;</td>
<td>&#x2713;</td>
<td>&#x2713;</td>
<td>&#x2713;</td>
<td>&#x2713;</td>
<td><bold>24.521</bold></td>
<td><bold>0.907</bold></td>
</tr>
</tbody>
</table>
</table-wrap>
 
<p><bold>1. Effectiveness of Topology-Preserving Isoline Routing.</bold> To test the routing algorithm, we separately examine the effects of hard-quantized iso-potential bins and stable sorting.</p>
<p>First, <bold>Model B (Continuous Potential)</bold> removes the potential quantization process and orders continuous floating-point potentials directly. Although <bold>Model B</bold> improves PSNR over the baseline from 23.552 to 23.957 dB, its SSIM decreases from 0.902 to 0.898. Since there are usually no identical potential values, the resulting sequence can frequently switch across distant spatial locations. This <italic>micro-fragmentation</italic> weakens local structure; therefore, discrete iso-potential layers are useful for preserving macroscopically continuous structures.</p>
<p>Second, <bold>Model C (Unstable Sort)</bold> uses isoline bins but employs a non-stable sorting method. Compared with the full model, PSNR decreases from 24.521 to 23.991 dB, and SSIM decreases from 0.907 to 0.904. Without stable sorting to preserve the local order within iso-potential strata, the resulting token sequence becomes more fragmented, forcing the network to compensate for more complex spatial discontinuities during reconstruction, thereby increasing the difficulty of image enhancement.</p>
<p><bold>2. Synergy of Potential Evolution and Physical Modulation.</bold> We introduce two coupled physics-inspired mechanisms in our architecture: Potential Field Evolution (<inline-formula id="ieqn-124"><mml:math id="mml-ieqn-124"><mml:mi mathvariant="normal">&#x0394;</mml:mi><mml:mi>P</mml:mi></mml:math></inline-formula>) for potential calibration and Transmission-Inspired Modulation (<inline-formula id="ieqn-125"><mml:math id="mml-ieqn-125"><mml:msub><mml:mi>W</mml:mi><mml:mrow><mml:mrow><mml:mi mathvariant="normal">m</mml:mi><mml:mi mathvariant="normal">o</mml:mi><mml:mi mathvariant="normal">d</mml:mi></mml:mrow></mml:mrow></mml:msub></mml:math></inline-formula>) for attenuation-aware feature reweighting and gradient propagation, as illustrated in <xref ref-type="fig" rid="fig-6">Fig. 6</xref>. To fully separate their responsibilities, we examine a series of progressive variants.</p>
<fig id="fig-6">
<label>Figure 6</label>
<caption>
<title>Visualization of the internal potential stream and modulation mechanism. From (<bold>a</bold>)&#x2013;(<bold>f</bold>): input image, geometric prior <inline-formula id="ieqn-126"><mml:math id="mml-ieqn-126"><mml:msub><mml:mi>P</mml:mi><mml:mn>0</mml:mn></mml:msub></mml:math></inline-formula>, evolved optical potential <inline-formula id="ieqn-127"><mml:math id="mml-ieqn-127"><mml:msub><mml:mi>P</mml:mi><mml:mrow><mml:mrow><mml:mi mathvariant="normal">n</mml:mi><mml:mi mathvariant="normal">e</mml:mi><mml:mi mathvariant="normal">w</mml:mi></mml:mrow></mml:mrow></mml:msub></mml:math></inline-formula>, evolution residual <inline-formula id="ieqn-128"><mml:math id="mml-ieqn-128"><mml:mi mathvariant="normal">&#x0394;</mml:mi><mml:mi>P</mml:mi></mml:math></inline-formula>, modulation gate <inline-formula id="ieqn-129"><mml:math id="mml-ieqn-129"><mml:msub><mml:mi>W</mml:mi><mml:mrow><mml:mrow><mml:mi mathvariant="normal">m</mml:mi><mml:mi mathvariant="normal">o</mml:mi><mml:mi mathvariant="normal">d</mml:mi></mml:mrow></mml:mrow></mml:msub></mml:math></inline-formula>, and restored image. The residual <inline-formula id="ieqn-130"><mml:math id="mml-ieqn-130"><mml:mi mathvariant="normal">&#x0394;</mml:mi><mml:mi>P</mml:mi></mml:math></inline-formula> calibrates <inline-formula id="ieqn-131"><mml:math id="mml-ieqn-131"><mml:msub><mml:mi>P</mml:mi><mml:mn>0</mml:mn></mml:msub></mml:math></inline-formula> into <inline-formula id="ieqn-132"><mml:math id="mml-ieqn-132"><mml:msub><mml:mi>P</mml:mi><mml:mrow><mml:mrow><mml:mi mathvariant="normal">n</mml:mi><mml:mi mathvariant="normal">e</mml:mi><mml:mi mathvariant="normal">w</mml:mi></mml:mrow></mml:mrow></mml:msub></mml:math></inline-formula>, while <inline-formula id="ieqn-133"><mml:math id="mml-ieqn-133"><mml:msub><mml:mi>W</mml:mi><mml:mrow><mml:mrow><mml:mi mathvariant="normal">m</mml:mi><mml:mi mathvariant="normal">o</mml:mi><mml:mi mathvariant="normal">d</mml:mi></mml:mrow></mml:mrow></mml:msub></mml:math></inline-formula> performs transmission-inspired spatially adaptive feature reweighting for scattering-dominant regions. In the <inline-formula id="ieqn-134"><mml:math id="mml-ieqn-134"><mml:mi mathvariant="normal">&#x0394;</mml:mi><mml:mi>P</mml:mi></mml:math></inline-formula> map, red and blue denote positive and negative corrections, respectively; brighter values in the <inline-formula id="ieqn-135"><mml:math id="mml-ieqn-135"><mml:msub><mml:mi>W</mml:mi><mml:mrow><mml:mrow><mml:mi mathvariant="normal">m</mml:mi><mml:mi mathvariant="normal">o</mml:mi><mml:mi mathvariant="normal">d</mml:mi></mml:mrow></mml:mrow></mml:msub></mml:math></inline-formula> map indicate larger gating responses.</title>
</caption>
<graphic mimetype="image" mime-subtype="tif" xlink:href="CMC_82357-fig-6.tif"/>
</fig>
<p>First, <bold>Model D (Routing Only)</bold> uses only a static geometric prior (<inline-formula id="ieqn-136"><mml:math id="mml-ieqn-136"><mml:msub><mml:mi>P</mml:mi><mml:mn>0</mml:mn></mml:msub></mml:math></inline-formula>) for Mamba sequencing. Using only the routing path, its PSNR reaches 23.899 dB, outperforming Model A. However, directly modulating features using this raw prior, as in <bold>Model E (Static Prior)</bold>, reduces PSNR to 23.682 dB. This <italic>negative transfer</italic> suggests that uncalibrated geometric distance should not be directly treated as underwater optical attenuation. Geometric depth and underwater optical degradation are correlated but not identical, and direct modulation may suppress useful foreground responses. Finally, <bold>Model F (Full Model)</bold> achieves the highest PSNR of 24.521 dB. This indicates that dynamic potential calibration is needed to reduce the domain gap of Model E. The <inline-formula id="ieqn-137"><mml:math id="mml-ieqn-137"><mml:msub><mml:mi>W</mml:mi><mml:mrow><mml:mrow><mml:mi mathvariant="normal">m</mml:mi><mml:mi mathvariant="normal">o</mml:mi><mml:mi mathvariant="normal">d</mml:mi></mml:mrow></mml:mrow></mml:msub></mml:math></inline-formula> gate serves two functions: it provides transmission-inspired spatial reweighting and establishes a continuous, differentiable optimization path around the non-differentiable <monospace>argsort</monospace> operation.</p>
<p><bold>3. Representation Capacity under Deep Physical Constraints.</bold> A concern with potential-guided modulation is that explicit spatial constraints may over-restrict deep feature representation. We therefore monitor the gate statistics of <inline-formula id="ieqn-138"><mml:math id="mml-ieqn-138"><mml:msub><mml:mi>W</mml:mi><mml:mrow><mml:mrow><mml:mi mathvariant="normal">m</mml:mi><mml:mi mathvariant="normal">o</mml:mi><mml:mi mathvariant="normal">d</mml:mi></mml:mrow></mml:mrow></mml:msub></mml:math></inline-formula> across hierarchical levels after training. If the modulation were harmful, the network could weaken this branch by driving the gate toward an identity-like response, i.e., <inline-formula id="ieqn-139"><mml:math id="mml-ieqn-139"><mml:msub><mml:mi>W</mml:mi><mml:mrow><mml:mrow><mml:mi mathvariant="normal">m</mml:mi><mml:mi mathvariant="normal">o</mml:mi><mml:mi mathvariant="normal">d</mml:mi></mml:mrow></mml:mrow></mml:msub><mml:mo>&#x2248;</mml:mo><mml:mn>1</mml:mn></mml:math></inline-formula>. Empirically, however, <inline-formula id="ieqn-140"><mml:math id="mml-ieqn-140"><mml:msub><mml:mi>W</mml:mi><mml:mrow><mml:mrow><mml:mi mathvariant="normal">m</mml:mi><mml:mi mathvariant="normal">o</mml:mi><mml:mi mathvariant="normal">d</mml:mi></mml:mrow></mml:mrow></mml:msub></mml:math></inline-formula> does not collapse to such an identity-like gate. Its mean gate values range from 0.428 to 0.521 across hierarchical levels, indicating that the modulation branch remains effective from shallow encoders to deep decoders and that underwater degradation remains relevant across multiple semantic scales rather than only at low-level layers.</p>
<p>This observation supports the division of roles in IG-Mamba. The routing branch organizes long-range dependencies along geometry-aware strata, the potential updater adapts the external prior to underwater restoration, and the modulation branch suppresses scattering-dominant feature responses without replacing the main restoration stream. Therefore, the performance drop of the Static Prior variant mainly arises from using an uncalibrated prior for modulation, rather than from insufficient representation capacity under physical guidance.</p>
</sec>
</sec>
<sec id="s5">
<label>5</label>
<title>Conclusion</title>
<p>This paper presents IG-Mamba, a physics-inspired framework that organizes Mamba state propagation with a depth-correlated potential prior. Topology-Preserving Isoline Scanning constructs geometry-aware sequences while preserving local row- and column-wise order within iso-potential strata. Potential Field Evolution further adapts the external geometry prior into a restoration-oriented optical potential for transmission-inspired feature reweighting. Experiments, degraded-prior stress tests, efficiency analysis, and bin-number sensitivity studies demonstrate that IG-Mamba provides an effective and practical geometry-aware SSM framework for underwater image restoration.</p>
</sec>
</body>
<back>
<ack>
<p>Not applicable.</p>
</ack>
<sec>
<title>Funding Statement</title>
<p>This work was supported in part by the National Natural Science Foundation of China (No. 62301105), in part by the Open Project of the State Key Laboratory of Ocean Sensing (No. OSKF-2025M04), in part by the Liaoning Provincial Science and Technology Plan Joint Program (Natural Science Foundation-General Program) (No. 2025-MSLH-113), in part by the Fundamental Research Funds for the Central Universities (No. 3132026240 and 3132025268), in part by the Fundamental Research Funds of the Liaoning Provincial Department of Education (No. LJ212510151021).</p>
</sec>
<sec>
<title>Author Contributions</title>
<p>The authors confirm contribution to the paper as follows: conceptualization, methodology, software, formal analysis, visualization, and writing&#x2014;original draft: Yiqiao Xiang; data curation and validation: Yiqiao Xiang, Ruijie Liu and Dehuan Zhang; writing&#x2014;review and editing: Jingchun Zhou, Ruijie Liu and Dehuan Zhang; supervision, project administration, and funding acquisition: Jingchun Zhou. All authors reviewed and approved the final version of the manuscript.</p>
</sec>
<sec sec-type="data-availability">
<title>Availability of Data and Materials</title>
<p>The authors confirm that the data supporting the findings of this study are available within the article. The public benchmark datasets used in this study are available from their original sources cited in the article.</p>
</sec>
<sec>
<title>Ethics Approval</title>
<p>Not applicable. This study did not involve human participants or animals.</p>
</sec>
<sec sec-type="COI-statement">
<title>Conflicts of Interest</title>
<p>The authors declare no conflicts of interest.</p>
</sec>
<ref-list content-type="authoryear">
<title>References</title>
<ref id="ref-1"><label>[1]</label><mixed-citation publication-type="journal"><person-group person-group-type="author"><string-name><surname>Song</surname> <given-names>W</given-names></string-name>, <string-name><surname>Wang</surname> <given-names>H</given-names></string-name>, <string-name><surname>Zhang</surname> <given-names>X</given-names></string-name>, <string-name><surname>Xia</surname> <given-names>J</given-names></string-name>, <string-name><surname>Liu</surname> <given-names>T</given-names></string-name>, <string-name><surname>Shi</surname> <given-names>Y</given-names></string-name></person-group>. <article-title>Deep-sea nodule mineral image segmentation algorithm based on Pix2PixHD</article-title>. <source>Comput Mater Contin</source>. <year>2022</year>;<volume>73</volume>(<issue>1</issue>):<fpage>1449</fpage>&#x2013;<lpage>62</lpage>. doi:<pub-id pub-id-type="doi">10.32604/cmc.2022.027213</pub-id>.</mixed-citation></ref>
<ref id="ref-2"><label>[2]</label><mixed-citation publication-type="journal"><person-group person-group-type="author"><string-name><surname>Li</surname> <given-names>C</given-names></string-name>, <string-name><surname>Wang</surname> <given-names>J</given-names></string-name>, <string-name><surname>Ding</surname> <given-names>C</given-names></string-name>, <string-name><surname>Ye</surname> <given-names>Z</given-names></string-name></person-group>. <article-title>AquaTree: deep reinforcement learning-driven Monte Carlo tree search for underwater image enhancement</article-title>. <source>Comput Mater Contin</source>. <year>2026</year>;<volume>86</volume>(<issue>3</issue>):<fpage>61</fpage>.</mixed-citation></ref>
<ref id="ref-3"><label>[3]</label><mixed-citation publication-type="journal"><person-group person-group-type="author"><string-name><surname>Zhou</surname> <given-names>J</given-names></string-name>, <string-name><surname>Zhang</surname> <given-names>D</given-names></string-name>, <string-name><surname>He</surname> <given-names>Z</given-names></string-name>, <string-name><surname>Gai</surname> <given-names>Q</given-names></string-name>, <string-name><surname>Jiang</surname> <given-names>Q</given-names></string-name></person-group>. <article-title>Multi-prior fusion transfer plugin for adapting in-air models to underwater image enhancement and detection</article-title>. <source>IEEE Trans Image Process</source>. <year>2025</year>;<volume>34</volume>:<fpage>7773</fpage>&#x2013;<lpage>85</lpage>. doi:<pub-id pub-id-type="doi">10.1109/tip.2025.3634001</pub-id>; <pub-id pub-id-type="pmid">41284421</pub-id></mixed-citation></ref>
<ref id="ref-4"><label>[4]</label><mixed-citation publication-type="journal"><person-group person-group-type="author"><string-name><surname>Yang</surname> <given-names>T</given-names></string-name>, <string-name><surname>Jia</surname> <given-names>S</given-names></string-name>, <string-name><surname>Ma</surname> <given-names>H</given-names></string-name></person-group>. <article-title>Research on the application of super resolution reconstruction algorithm for underwater image</article-title>. <source>Comput Mater Contin</source>. <year>2020</year>;<volume>62</volume>(<issue>3</issue>):<fpage>1249</fpage>. doi:<pub-id pub-id-type="doi">10.32604/cmc.2020.05777</pub-id>.</mixed-citation></ref>
<ref id="ref-5"><label>[5]</label><mixed-citation publication-type="journal"><person-group person-group-type="author"><string-name><surname>Zhu</surname> <given-names>X</given-names></string-name>, <string-name><surname>Zhang</surname> <given-names>J</given-names></string-name>, <string-name><surname>Zeng</surname> <given-names>H</given-names></string-name></person-group>. <article-title>FENet: underwater image enhancement via frequency domain enhancement and edge-guided refinement</article-title>. <source>Comput Mater Contin</source>. <year>2026</year>;<volume>86</volume>(<issue>2</issue>):<fpage>1</fpage>.</mixed-citation></ref>
<ref id="ref-6"><label>[6]</label><mixed-citation publication-type="journal"><person-group person-group-type="author"><string-name><surname>Zhou</surname> <given-names>J</given-names></string-name>, <string-name><surname>He</surname> <given-names>Z</given-names></string-name>, <string-name><surname>Zhang</surname> <given-names>D</given-names></string-name>, <string-name><surname>Liu</surname> <given-names>S</given-names></string-name>, <string-name><surname>Fu</surname> <given-names>X</given-names></string-name>, <string-name><surname>Li</surname> <given-names>X</given-names></string-name></person-group>. <article-title>Spatial residual for underwater object detection</article-title>. <source>IEEE Trans Pattern Anal Mach Intell</source>. <year>2025</year>;<volume>47</volume>(<issue>6</issue>):<fpage>4996</fpage>&#x2013;<lpage>5013</lpage>. doi:<pub-id pub-id-type="doi">10.1109/tpami.2025.3548652</pub-id>; <pub-id pub-id-type="pmid">40048345</pub-id></mixed-citation></ref>
<ref id="ref-7"><label>[7]</label><mixed-citation publication-type="conf-proc"><person-group person-group-type="author"><string-name><surname>Gu</surname> <given-names>A</given-names></string-name>, <string-name><surname>Dao</surname> <given-names>T</given-names></string-name></person-group>. <article-title>Mamba: linear-time sequence modeling with selective state spaces</article-title>. In: <conf-name>Proceedings of the First Conference on Language Modeling; 2024 Oct 7&#x2013;10; Philadelphia, PA, USA</conf-name>.</mixed-citation></ref>
<ref id="ref-8"><label>[8]</label><mixed-citation publication-type="conf-proc"><person-group person-group-type="author"><string-name><surname>Akkaynak</surname> <given-names>D</given-names></string-name>, <string-name><surname>Treibitz</surname> <given-names>T</given-names></string-name></person-group>. <article-title>Sea-thru: a method for removing water from underwater images</article-title>. In: <conf-name>Proceedings of the IEEE/CVF Conference on Computer Vision and Pattern Recognition; 2019 Jun 15&#x2013;20; Long Beach, CA, USA</conf-name>. p. <fpage>1682</fpage>&#x2013;<lpage>91</lpage>.</mixed-citation></ref>
<ref id="ref-9"><label>[9]</label><mixed-citation publication-type="journal"><person-group person-group-type="author"><string-name><surname>Jaffe</surname> <given-names>JS</given-names></string-name></person-group>. <article-title>Computer modeling and the design of optimal underwater imaging systems</article-title>. <source>IEEE J Ocean Eng</source>. <year>1990</year>;<volume>15</volume>(<issue>2</issue>):<fpage>101</fpage>&#x2013;<lpage>11</lpage>. doi:<pub-id pub-id-type="doi">10.1109/48.50695</pub-id>.</mixed-citation></ref>
<ref id="ref-10"><label>[10]</label><mixed-citation publication-type="journal"><person-group person-group-type="author"><string-name><surname>He</surname> <given-names>K</given-names></string-name>, <string-name><surname>Sun</surname> <given-names>J</given-names></string-name>, <string-name><surname>Tang</surname> <given-names>X</given-names></string-name></person-group>. <article-title>Single image haze removal using dark channel prior</article-title>. <source>IEEE Trans Pattern Anal Mach Intell</source>. <year>2010</year>;<volume>33</volume>(<issue>12</issue>):<fpage>2341</fpage>&#x2013;<lpage>53</lpage>. doi:<pub-id pub-id-type="doi">10.1109/tpami.2010.168</pub-id>; <pub-id pub-id-type="pmid">20820075</pub-id></mixed-citation></ref>
<ref id="ref-11"><label>[11]</label><mixed-citation publication-type="journal"><person-group person-group-type="author"><string-name><surname>Li</surname> <given-names>C</given-names></string-name>, <string-name><surname>Guo</surname> <given-names>C</given-names></string-name>, <string-name><surname>Ren</surname> <given-names>W</given-names></string-name>, <string-name><surname>Cong</surname> <given-names>R</given-names></string-name>, <string-name><surname>Hou</surname> <given-names>J</given-names></string-name>, <string-name><surname>Kwong</surname> <given-names>S</given-names></string-name>, <etal>et al</etal></person-group>. <article-title>An underwater image enhancement benchmark dataset and beyond</article-title>. <source>IEEE Trans Image Process</source>. <year>2019</year>;<volume>29</volume>:<fpage>4376</fpage>&#x2013;<lpage>89</lpage>. doi:<pub-id pub-id-type="doi">10.1109/TIP.2019.2955241</pub-id>; <pub-id pub-id-type="pmid">31796402</pub-id></mixed-citation></ref>
<ref id="ref-12"><label>[12]</label><mixed-citation publication-type="journal"><person-group person-group-type="author"><string-name><surname>Li</surname> <given-names>C</given-names></string-name>, <string-name><surname>Anwar</surname> <given-names>S</given-names></string-name>, <string-name><surname>Porikli</surname> <given-names>F</given-names></string-name></person-group>. <article-title>Underwater scene prior inspired deep underwater image and video enhancement</article-title>. <source>Pattern Recognit</source>. <year>2020</year>;<volume>98</volume>(<issue>1</issue>):<fpage>107038</fpage>. doi:<pub-id pub-id-type="doi">10.1016/j.patcog.2019.107038</pub-id>.</mixed-citation></ref>
<ref id="ref-13"><label>[13]</label><mixed-citation publication-type="journal"><person-group person-group-type="author"><string-name><surname>Li</surname> <given-names>C</given-names></string-name>, <string-name><surname>Anwar</surname> <given-names>S</given-names></string-name>, <string-name><surname>Hou</surname> <given-names>J</given-names></string-name>, <string-name><surname>Cong</surname> <given-names>R</given-names></string-name>, <string-name><surname>Guo</surname> <given-names>C</given-names></string-name>, <string-name><surname>Ren</surname> <given-names>W</given-names></string-name></person-group>. <article-title>Underwater image enhancement via medium transmission-guided multi-color space embedding</article-title>. <source>IEEE Trans Image Process</source>. <year>2021</year>;<volume>30</volume>:<fpage>4985</fpage>&#x2013;<lpage>5000</lpage>. doi:<pub-id pub-id-type="doi">10.1109/tip.2021.3076367</pub-id>; <pub-id pub-id-type="pmid">33961554</pub-id></mixed-citation></ref>
<ref id="ref-14"><label>[14]</label><mixed-citation publication-type="journal"><person-group person-group-type="author"><string-name><surname>Peng</surname> <given-names>L</given-names></string-name>, <string-name><surname>Zhu</surname> <given-names>C</given-names></string-name>, <string-name><surname>Bian</surname> <given-names>L</given-names></string-name></person-group>. <article-title>U-shape transformer for underwater image enhancement</article-title>. <source>IEEE Trans Image Process</source>. <year>2023</year>;<volume>32</volume>(<issue>2</issue>):<fpage>3066</fpage>&#x2013;<lpage>79</lpage>. doi:<pub-id pub-id-type="doi">10.1109/tip.2023.3276332</pub-id>; <pub-id pub-id-type="pmid">37200123</pub-id></mixed-citation></ref>
<ref id="ref-15"><label>[15]</label><mixed-citation publication-type="journal"><person-group person-group-type="author"><string-name><surname>Liu</surname> <given-names>Y</given-names></string-name>, <string-name><surname>Tian</surname> <given-names>Y</given-names></string-name>, <string-name><surname>Zhao</surname> <given-names>Y</given-names></string-name>, <string-name><surname>Yu</surname> <given-names>H</given-names></string-name>, <string-name><surname>Xie</surname> <given-names>L</given-names></string-name>, <string-name><surname>Wang</surname> <given-names>Y</given-names></string-name>, <etal>et al</etal></person-group>. <article-title>Vmamba: visual state space model</article-title>. <source>Adv Neural Inf Process Syst</source>. <year>2024</year>;<volume>37</volume>:<fpage>103031</fpage>&#x2013;<lpage>63</lpage>.</mixed-citation></ref>
<ref id="ref-16"><label>[16]</label><mixed-citation publication-type="journal"><person-group person-group-type="author"><string-name><surname>Yang</surname> <given-names>G</given-names></string-name>, <string-name><surname>Kang</surname> <given-names>G</given-names></string-name>, <string-name><surname>Lee</surname> <given-names>J</given-names></string-name>, <string-name><surname>Cho</surname> <given-names>Y</given-names></string-name></person-group>. <article-title>Joint-ID: transformer-based joint image enhancement and depth estimation for underwater environments</article-title>. <source>IEEE Sens J</source>. <year>2023</year>;<volume>24</volume>(<issue>3</issue>):<fpage>3113</fpage>&#x2013;<lpage>22</lpage>.</mixed-citation></ref>
<ref id="ref-17"><label>[17]</label><mixed-citation publication-type="other"><person-group person-group-type="author"><string-name><surname>Du</surname> <given-names>D</given-names></string-name>, <string-name><surname>Si</surname> <given-names>L</given-names></string-name>, <string-name><surname>Xu</surname> <given-names>F</given-names></string-name>, <string-name><surname>Niu</surname> <given-names>J</given-names></string-name>, <string-name><surname>Sun</surname> <given-names>F</given-names></string-name></person-group>. <article-title>A physical model-guided framework for underwater image enhancement and depth estimation</article-title>. <comment>arXiv:2407.04230. 2024</comment>.</mixed-citation></ref>
<ref id="ref-18"><label>[18]</label><mixed-citation publication-type="journal"><person-group person-group-type="author"><string-name><surname>Huang</surname> <given-names>Z</given-names></string-name>, <string-name><surname>Wang</surname> <given-names>X</given-names></string-name>, <string-name><surname>Xu</surname> <given-names>C</given-names></string-name>, <string-name><surname>Li</surname> <given-names>J</given-names></string-name>, <string-name><surname>Feng</surname> <given-names>L</given-names></string-name></person-group>. <article-title>Underwater variable zoom: depth-guided perception network for underwater image enhancement</article-title>. <source>Expert Syst Appl</source>. <year>2025</year>;<volume>259</volume>:<fpage>125350</fpage>.</mixed-citation></ref>
<ref id="ref-19"><label>[19]</label><mixed-citation publication-type="journal"><person-group person-group-type="author"><string-name><surname>Yang</surname> <given-names>L</given-names></string-name>, <string-name><surname>Kang</surname> <given-names>B</given-names></string-name>, <string-name><surname>Huang</surname> <given-names>Z</given-names></string-name>, <string-name><surname>Zhao</surname> <given-names>Z</given-names></string-name>, <string-name><surname>Xu</surname> <given-names>X</given-names></string-name>, <string-name><surname>Feng</surname> <given-names>J</given-names></string-name>, <etal>et al</etal></person-group>. <article-title>Depth anything v2</article-title>. <source>Adv Neural Inf Process Syst</source>. <year>2024</year>;<volume>37</volume>:<fpage>21875</fpage>&#x2013;<lpage>911</lpage>. doi:<pub-id pub-id-type="doi">10.52202/079017-0688</pub-id>.</mixed-citation></ref>
<ref id="ref-20"><label>[20]</label><mixed-citation publication-type="other"><person-group person-group-type="author"><string-name><surname>Li</surname> <given-names>H</given-names></string-name>, <string-name><surname>Li</surname> <given-names>J</given-names></string-name>, <string-name><surname>Wang</surname> <given-names>W</given-names></string-name></person-group>. <article-title>A fusion adversarial underwater image enhancement network with a public test dataset</article-title>. <comment>arXiv:1906.06819. 2019</comment>.</mixed-citation></ref>
<ref id="ref-21"><label>[21]</label><mixed-citation publication-type="journal"><person-group person-group-type="author"><string-name><surname>Liu</surname> <given-names>R</given-names></string-name>, <string-name><surname>Fan</surname> <given-names>X</given-names></string-name>, <string-name><surname>Zhu</surname> <given-names>M</given-names></string-name>, <string-name><surname>Hou</surname> <given-names>M</given-names></string-name>, <string-name><surname>Luo</surname> <given-names>Z</given-names></string-name></person-group>. <article-title>Real-world underwater enhancement: challenges, benchmarks, and solutions under natural light</article-title>. <source>IEEE Trans Circuits Syst Video Technol</source>. <year>2020</year>;<volume>30</volume>(<issue>12</issue>):<fpage>4861</fpage>&#x2013;<lpage>75</lpage>.</mixed-citation></ref>
<ref id="ref-22"><label>[22]</label><mixed-citation publication-type="journal"><person-group person-group-type="author"><string-name><surname>Song</surname> <given-names>W</given-names></string-name>, <string-name><surname>Wang</surname> <given-names>Y</given-names></string-name>, <string-name><surname>Huang</surname> <given-names>D</given-names></string-name>, <string-name><surname>Liotta</surname> <given-names>A</given-names></string-name>, <string-name><surname>Perra</surname> <given-names>C</given-names></string-name></person-group>. <article-title>Enhancement of underwater images with statistical model of background light and optimization of transmission map</article-title>. <source>IEEE Trans Broadcast</source>. <year>2020</year>;<volume>66</volume>(<issue>1</issue>):<fpage>153</fpage>&#x2013;<lpage>69</lpage>. doi:<pub-id pub-id-type="doi">10.1109/tbc.2019.2960942</pub-id>.</mixed-citation></ref>
<ref id="ref-23"><label>[23]</label><mixed-citation publication-type="other"><person-group person-group-type="author"><string-name><surname>Islam</surname> <given-names>MJ</given-names></string-name>, <string-name><surname>Luo</surname> <given-names>P</given-names></string-name>, <string-name><surname>Sattar</surname> <given-names>J</given-names></string-name></person-group>. <article-title>Simultaneous enhancement and super-resolution of underwater imagery for improved visual perception</article-title>. <comment>arXiv:2002.01155. 2020</comment>.</mixed-citation></ref>
<ref id="ref-24"><label>[24]</label><mixed-citation publication-type="journal"><person-group person-group-type="author"><string-name><surname>Zhou</surname> <given-names>J</given-names></string-name>, <string-name><surname>Sun</surname> <given-names>J</given-names></string-name>, <string-name><surname>Li</surname> <given-names>C</given-names></string-name>, <string-name><surname>Jiang</surname> <given-names>Q</given-names></string-name>, <string-name><surname>Zhou</surname> <given-names>M</given-names></string-name>, <string-name><surname>Lam</surname> <given-names>KM</given-names></string-name>, <etal>et al</etal></person-group>. <article-title>HCLR-Net: hybrid contrastive learning regularization with locally randomized perturbation for underwater image enhancement</article-title>. <source>Int J Comput Vis</source>. <year>2024</year>;<volume>132</volume>(<issue>10</issue>):<fpage>4132</fpage>&#x2013;<lpage>56</lpage>.</mixed-citation></ref>
<ref id="ref-25"><label>[25]</label><mixed-citation publication-type="conf-proc"><person-group person-group-type="author"><string-name><surname>Korhonen</surname> <given-names>J</given-names></string-name>, <string-name><surname>You</surname> <given-names>J</given-names></string-name></person-group>. <article-title>Peak signal-to-noise ratio revisited: is simple beautiful?</article-title>. In: <conf-name>2012 Fourth International Workshop on Quality of Multimedia Experience; 2012 Jul 5&#x2013;7; Melbourne, VIC, Australia</conf-name>. p. <fpage>37</fpage>&#x2013;<lpage>8</lpage>.</mixed-citation></ref>
<ref id="ref-26"><label>[26]</label><mixed-citation publication-type="journal"><person-group person-group-type="author"><string-name><surname>Wang</surname> <given-names>Z</given-names></string-name>, <string-name><surname>Bovik</surname> <given-names>AC</given-names></string-name>, <string-name><surname>Sheikh</surname> <given-names>HR</given-names></string-name>, <string-name><surname>Simoncelli</surname> <given-names>EP</given-names></string-name></person-group>. <article-title>Image quality assessment: from error visibility to structural similarity</article-title>. <source>IEEE Trans Image Process</source>. <year>2004</year>;<volume>13</volume>(<issue>4</issue>):<fpage>600</fpage>&#x2013;<lpage>12</lpage>. doi:<pub-id pub-id-type="doi">10.1109/tip.2003.819861</pub-id>; <pub-id pub-id-type="pmid">15376593</pub-id></mixed-citation></ref>
<ref id="ref-27"><label>[27]</label><mixed-citation publication-type="journal"><person-group person-group-type="author"><string-name><surname>Yang</surname> <given-names>M</given-names></string-name>, <string-name><surname>Sowmya</surname> <given-names>A</given-names></string-name></person-group>. <article-title>An underwater color image quality evaluation metric</article-title>. <source>IEEE Trans Image Process</source>. <year>2015</year>;<volume>24</volume>(<issue>12</issue>):<fpage>6062</fpage>&#x2013;<lpage>71</lpage>. doi:<pub-id pub-id-type="doi">10.1109/tip.2015.2491020</pub-id>; <pub-id pub-id-type="pmid">26513783</pub-id></mixed-citation></ref>
<ref id="ref-28"><label>[28]</label><mixed-citation publication-type="journal"><person-group person-group-type="author"><string-name><surname>Guo</surname> <given-names>C</given-names></string-name>, <string-name><surname>Wu</surname> <given-names>R</given-names></string-name>, <string-name><surname>Jin</surname> <given-names>X</given-names></string-name>, <string-name><surname>Han</surname> <given-names>L</given-names></string-name>, <string-name><surname>Zhang</surname> <given-names>W</given-names></string-name>, <string-name><surname>Chai</surname> <given-names>Z</given-names></string-name>, <etal>et al</etal></person-group>. <article-title>Underwater ranker: learn which is better and how to be better</article-title>. <source>Proc AAAI Conf Artif Intell</source>. <year>2023</year>;<volume>37</volume>:<fpage>702</fpage>&#x2013;<lpage>9</lpage>.</mixed-citation></ref>
<ref id="ref-29"><label>[29]</label><mixed-citation publication-type="journal"><person-group person-group-type="author"><string-name><surname>Kang</surname> <given-names>Y</given-names></string-name>, <string-name><surname>Jiang</surname> <given-names>Q</given-names></string-name>, <string-name><surname>Li</surname> <given-names>C</given-names></string-name>, <string-name><surname>Ren</surname> <given-names>W</given-names></string-name>, <string-name><surname>Liu</surname> <given-names>H</given-names></string-name>, <string-name><surname>Wang</surname> <given-names>P</given-names></string-name></person-group>. <article-title>A perception-aware decomposition and fusion framework for underwater image enhancement</article-title>. <source>IEEE Trans Circuits Syst Video Technol</source>. <year>2022</year>;<volume>33</volume>(<issue>3</issue>):<fpage>988</fpage>&#x2013;<lpage>1002</lpage>. doi:<pub-id pub-id-type="doi">10.1109/tcsvt.2022.3208100</pub-id>.</mixed-citation></ref>
<ref id="ref-30"><label>[30]</label><mixed-citation publication-type="conf-proc"><person-group person-group-type="author"><string-name><surname>Zhou</surname> <given-names>S</given-names></string-name>, <string-name><surname>Chen</surname> <given-names>D</given-names></string-name>, <string-name><surname>Pan</surname> <given-names>J</given-names></string-name>, <string-name><surname>Shi</surname> <given-names>J</given-names></string-name>, <string-name><surname>Yang</surname> <given-names>J</given-names></string-name></person-group>. <article-title>Adapt or perish: adaptive sparse transformer with attentive feature refinement for image restoration</article-title>. In: <conf-name>Proceedings of the IEEE/CVF Conference on Computer Vision and Pattern Recognition (CVPR); 2024 Jun 16&#x2013;22; Seattle, WA, USA</conf-name>. p. <fpage>2952</fpage>&#x2013;<lpage>63</lpage>.</mixed-citation></ref>
<ref id="ref-31"><label>[31]</label><mixed-citation publication-type="journal"><person-group person-group-type="author"><string-name><surname>Zou</surname> <given-names>W</given-names></string-name>, <string-name><surname>Gao</surname> <given-names>H</given-names></string-name>, <string-name><surname>Ye</surname> <given-names>T</given-names></string-name>, <string-name><surname>Chen</surname> <given-names>L</given-names></string-name>, <string-name><surname>Yang</surname> <given-names>W</given-names></string-name>, <string-name><surname>Huang</surname> <given-names>S</given-names></string-name>, <etal>et al</etal></person-group>. <article-title>VQCNIR: clearer night image restoration with vector-quantized codebook</article-title>. <source>Proc AAAI Conf Artif Intell</source>. <year>2024</year>;<volume>38</volume>(<issue>7</issue>):<fpage>7873</fpage>&#x2013;<lpage>81</lpage>. doi:<pub-id pub-id-type="doi">10.1609/aaai.v38i7.28623</pub-id>.</mixed-citation></ref>
<ref id="ref-32"><label>[32]</label><mixed-citation publication-type="journal"><person-group person-group-type="author"><string-name><surname>Luan</surname> <given-names>X</given-names></string-name>, <string-name><surname>Fan</surname> <given-names>H</given-names></string-name>, <string-name><surname>Wang</surname> <given-names>Q</given-names></string-name>, <string-name><surname>Yang</surname> <given-names>N</given-names></string-name>, <string-name><surname>Liu</surname> <given-names>S</given-names></string-name>, <string-name><surname>Li</surname> <given-names>X</given-names></string-name>, <etal>et al</etal></person-group>. <article-title>FMambaIR: a hybrid state space model and frequency domain for image restoration</article-title>. <source>IEEE Trans Geosci Remote Sens</source>. <year>2025</year>;<volume>63</volume>:<fpage>4201614</fpage>.</mixed-citation></ref>
<ref id="ref-33"><label>[33]</label><mixed-citation publication-type="journal"><person-group person-group-type="author"><string-name><surname>Saleh</surname> <given-names>A</given-names></string-name>, <string-name><surname>Sheaves</surname> <given-names>M</given-names></string-name>, <string-name><surname>Jerry</surname> <given-names>D</given-names></string-name>, <string-name><surname>Azghadi</surname> <given-names>MR</given-names></string-name></person-group>. <article-title>Adaptive deep learning framework for robust unsupervised underwater image enhancement</article-title>. <source>Expert Syst Appl</source>. <year>2025</year>;<volume>268</volume>(<issue>1</issue>):<fpage>126314</fpage>. doi:<pub-id pub-id-type="doi">10.1016/j.eswa.2024.126314</pub-id>.</mixed-citation></ref>
<ref id="ref-34"><label>[34]</label><mixed-citation publication-type="journal"><person-group person-group-type="author"><string-name><surname>Zhang</surname> <given-names>H</given-names></string-name>, <string-name><surname>Xu</surname> <given-names>H</given-names></string-name>, <string-name><surname>Yu</surname> <given-names>X</given-names></string-name>, <string-name><surname>Zhang</surname> <given-names>X</given-names></string-name>, <string-name><surname>Gao</surname> <given-names>X</given-names></string-name>, <string-name><surname>Wu</surname> <given-names>C</given-names></string-name></person-group>. <article-title>CDF-UIE: leveraging cross-domain fusion for underwater image enhancement</article-title>. <source>IEEE Trans Geosci Remote Sens</source>. <year>2025</year>;<volume>63</volume>:<fpage>4203715</fpage>. doi:<pub-id pub-id-type="doi">10.1109/tgrs.2025.3553557</pub-id>.</mixed-citation></ref>
</ref-list>
</back></article>









