REVIEW 4 major objections 4 minor 3 cited by
ProxelGen: Generating Proteins as 3D Densities
T0 review · 4 major / 4 minor · reviewed 2026-08-15 · deepseek-v4-flash
Pith's one-line read This paper claims protein structures can be generated directly as multi-channel 3D density grids, producing samples more novel and better matched to the training distribution than atomistic point-cloud models at comparable designability.
desk verdict The proxel representation and its spatial conditioning are the real contributions; the unconditional designability claim is currently entangled with a Proteina refinement step not applied to baselines. read the letter →
The pith
A machine-rendered reading of the paper's core claim, the machinery that carries it, and where it could break.
The reading
What carries the argument
The load-bearing object is the 'proxel' representation: a 7-channel 3D voxel grid in which three channels sample Gaussian-smoothed densities around backbone C, Cα, and N atoms, one channel encodes bond midpoints, and three channels form a 'chain flow' vector field pointing along successive Cα–Cα vectors from the N to the C terminus. That chain-flow channel is the mechanism that carries ordering information, making it possible to thread a single amino-acid chain through the generated density. On top of this representation, ProxelGen compresses 32×32×32 proxel arrays with a 3D CNN VAE into a latent space (512× spatial compression), trains a stochastic-interpolant flow model to generate latents, and fine-tunes an atomistic flow decoder to convert generated latents back to backbone coordinates. Because conditioning inputs are also voxel grids, spatial constraints can be injected by channel-wise concatenation in latent space.
What would settle it
Decode the generated proxels with the small coordinate decoder and score designability without the large refinement pass; if designability falls to near zero, the density representation itself is not carrying the reported result. Separately, count the connected components of the generated chain-flow field before decoding: if most samples split or merge, the ordering channel is not doing the threading job the method requires.
Extended reading notes
Core claim
The paper's central claim is that a protein structure can be generated as a 3D density rather than as an atomistic point cloud, and that this alternative representation is not merely viable but advantageous. ProxelGen encodes proteins as 'proxels'—multi-channel voxel arrays sampled from Gaussian-smoothed densities around the backbone atoms, plus a bond channel and a vector field that traces the chain from N- to C-terminus—and learns a 3D-convolutional VAE whose latent space is generated by a flow model. On unconditional generation, ProxelGen reports higher novelty, better FID against the training distribution, and designability at roughly native levels when compared with a leading atomistic flow model, and in motif scaffolding it reports 12 unique successful designs for a four-segment motif where every tested baseline finds one. The paper also demonstrates that spatial conditioning follows for free from the representation: masked regions and arbitrary voxelized shapes can be concatenated as inputs, enabling inpainting and shape-conditioned generation without prescribing protein length or the placement of motif segments.
Load-bearing premise
The argument assumes that the generated proxels—not the pretrained coordinate decoder and refinement model—are what make the sampled structures designable, and that the chain-flow channel reliably encodes a single connected chain; if either fails, the headline comparisons describe the whole pipeline rather than the density representation.
Editorial extensions
If this is right
- Protein length no longer has to be fixed before sampling: the same voxel grid can represent different numbers of residues, so generation and inpainting can produce whatever size the conditioning shape supports.
- Spatial tasks that require awkward bookkeeping in atomistic models—motif scaffolding, masked inpainting, shape-conditioned design—become native operations of concatenating or masking voxel channels.
- If the 1BCF and 1QJG results generalize, density-based sampling will be the preferred tool for scaffolding multi-segment motifs, where sequence-anchored methods effectively fail.
- The fixed grid lets protein generation borrow mature 3D CNN and latent diffusion infrastructure from image generation, and the representation itself scales with grid resolution rather than chain length.
Reading between the lines
- A decisive attribution test would be to score designability of decoded structures before the 200M-parameter refinement pass; if designability collapses without refinement, the headline result belongs to the decoder/refinement pipeline rather than to the density representation.
- If chain-flow connectivity is enforced as a training objective, density-based models should extend to multi-chain complexes and to conditioning on experimental density maps, where a single connected chain is not the right prior; the paper's current chain flow assumes one chain and the authors note it often splits or merges.
- The shape-conditioning setup suggests a direct application the paper leaves untested: conditioning on low-resolution experimental envelopes rather than shapes derived from known structures, using the same shape-adherence metrics.
- The fixed-resolution proxel grid also points toward a coarse-to-fine generation scheme—generate a low-resolution shape, then sample higher-resolution refinements—which atomistic representations cannot express naturally.
Signed reviews
Editorial analysis
A structured set of objections, weighed in public.
Referee Report
Summary. The paper introduces ProxelGen, a generative model of protein structure built on a voxelized 3D density representation called proxels. A protein is encoded as a multi-channel voxel grid containing backbone-atom Gaussian densities, a bond channel, and a three-channel chain-flow vector field. ProxelGen consists of a 3D CNN VAE that compresses proxels into a fixed-size latent space, a latent flow model trained to generate these latents, and a fine-tuned Proteina-based coordinate decoder that maps generated proxels back to atomistic backbones. The paper claims that on unconditional generation ProxelGen achieves higher novelty, better FID, and designability comparable to the training set, and that its spatial conditioning enables competitive motif scaffolding and shape-conditional generation. The central claim is that density-based protein generation is a viable alternative to atomistic point-cloud representations.
Significance. If the central claim holds, the paper opens a useful new axis in protein structure generation: representing proteins as voxelized densities rather than atomistic coordinates or frames. This representation naturally supports fixed-size latents, convolutional architectures, inpainting by spatial masking, and shape conditioning, and it connects protein structure generation to the mature literature on voxel-based generative modeling. The paper also provides useful validation of a self-supervised proxel embedding (ProxCLR) for FID-style evaluation, with sanity checks against perturbations and cluster removal. The use of external oracles (TM-align, ProteinMPNN/ESMFold self-consistency, FoldSeek diversity) is a strength, since the headline quantities are not defined purely by the model itself. However, the main empirical claims are currently not anchored tightly enough: the designability comparison is confounded by a large pretrained refinement model, the FID metric is ambiguously defined, and the reported differences lack statistical error bars. These issues are fixable and do not invalidate the core idea, but they must be addressed before the claims can be accepted.
major comments (4)
- [Section 3.4, Appendix B.1, Table 1] The central claim that ProxelGen achieves 'the same level of designability as the training set' is evaluated after a renoise-denoise refinement step using an unconditional 200M Proteina model at t=0.8, which is not applied to the Proteina baselines. The paper states that this refinement 'vastly improved' designability while moving structures by about 3.35 Å RMSD. Because the decoder and refiner are large pretrained atomistic priors that are not conditioned on the generated proxels, the reported designability may characterize the Proteina-assisted pipeline rather than the density representation. The Native Proxels row (designability 50.39) being close to ProxelGen's 53.13 strengthens this concern. A control in which random or scrambled proxels are passed through the same decoding and refinement pipeline is needed to attribute the observed designability to the generative model rather than to the decoder/refiner.
- [Section 4.2, Table 1] All unconditional metrics are computed on 256 samples with no error bars, no multiple seeds, and no statistical significance testing. Several headline differences are small (FID 6.05 vs 7.25 for Proteina 400M (H); novelty 0.73 vs 0.69; designability 53.13 vs 45.70), and the reader cannot tell whether these gaps are meaningful. Please report confidence intervals, multiple seeds, or a statistical comparison for the key metrics in Table 1.
- [Section 3.5, Section 4.2, Appendix B.1] The FID reported in Table 1 is not clearly defined. The text states that 'the FID of the generated proxels themselves, before decoding back to atomic coordinates, is 6.81,' but Table 1 reports FID=6.05 for ProxelGen. Since the refinement step affects both designability and FID, it is unclear whether Table 1's FID is computed on raw generated proxels via ProxCLR, on decoded structures via ProteinFID, or on proxelized versions of decoded and refined structures. The comparison with Proteina baselines is only meaningful if the same embedding and the same pre/post-processing are used for all methods. Please define the FID variant precisely and report pre-refinement and post-refinement values separately.
- [Section 6] The paper admits that many generated proxels do not form a single connected chain, that the chain flow may split or merge, and that the coordinate decoder often introduces unnatural kinks when threading a chain through such proxels. This is a load-bearing limitation for the claim that the density representation carries sufficient ordering information for full-chain generation. Please quantify how often chain connectivity fails in the generated proxels, and report how many of the designable samples in Table 1 and the successful scaffolding designs in Table 2 required decoder compensation for disconnected chain flow. Without this quantification, the reported results describe the combined proxel-plus-decoder pipeline rather than the density representation alone.
minor comments (4)
- [Appendix A.2] In the sentence describing the contrastive learning details, 'stochasic' should be 'stochastic'.
- [Section 3.1, Equation (6)] The chain flow vector field is defined using vectors v_i for i=1,...,N-1, but the summation in Equation (6) runs over i=1,...,N; the indexing for the terminal C-alpha position should be clarified.
- [Section 3.5 and elsewhere] The acronym 'FID' is used for both the ProxCLR-based proxel FID and the more general Fréchet Inception Distance naming convention; using a distinct term such as 'ProxFID' would avoid confusion with FID computed on atomistic structures.
- [Table 2] The merged length rows in the upper half of Table 2 are difficult to parse, especially because ProxelGen reports a single number per task while prespecified methods report one number per length range; the caption should explain the row structure more explicitly.
Circularity Check
No circularity: ProxelGen's representation, losses, and evaluations are externally anchored; the disclosed Proteina refinement weakens attribution but does not make any claimed result definitionally equivalent to its inputs.
full rationale
The derivation chain is self-contained rather than circular. The proxel representation is an explicit construction from backbone coordinates (Eqs. 1-6), the VAE and flow objectives are standard reconstruction and velocity-matching losses (Eqs. 7-8), and the headline claims are evaluated against external oracles: ProteinMPNN+ESMFold self-consistency RMSD for designability, FoldSeek clustering for diversity, TM-align for novelty, and external baselines (RFDiffusion, Genie2, FrameFlow, Proteina) for scaffolding. The FID metric uses a self-supervised ProxCLR embedding trained by the authors, but App. C validates it behaviorally against perturbations, FoldSeek clusters, and CATH hierarchy removals, so the metric is not defined into existence by the model being scored. The most serious caveat is App. B.1: the reported designability is obtained after renoising and denoising decoded structures with a fixed pretrained 200M Proteina model at t=0.8, which the paper says 'vastly improved' designability; this is a genuine attribution/control concern for the claim that the proxel representation alone yields designable structures, and a scrambled-proxel control would strengthen the paper. However, that is an experimental confound, not a circular reduction: the refiner is a fixed pretrained prior rather than a parameter fitted to the reported metric, and the equations do not reduce to their inputs. The admitted chain-connectivity failure in Sec. 6 likewise limits the method's practical validity but does not indicate circularity. No load-bearing step is carried by self-citation alone, and the central representation claim has independent empirical content.
Assumptions & free parameters
free parameters (8)
- Proxel grid spacing =
1.5 Å
- Gaussian kernel standard deviation =
1 Å
- Gaussian cutoff radius =
4.5 Å
- VAE KL weight beta =
1e-6
- Latent downsampling factor =
512x total compression, f not stated
- Flow velocity step-size multiplier =
1.5
- Timestep sampling mixture =
uniform [0,1] mixed with normal centered at t=0.2, sigma=0.1
- Structure refinement noise level =
t=0.8
assumptions (6)
- domain assumption Gaussian-smoothed atom channels plus chain flow capture enough information to recover designable backbones.
- domain assumption PCA axis alignment gives a canonical frame so non-equivariant 3D convolutions can model the distribution.
- domain assumption The chain-flow vector field uniquely determines chain order.
- domain assumption ProxCLR SimCLR embeddings yield a meaningful metric for FID.
- domain assumption ProteinMPNN plus ESMFold self-consistency is a valid designability oracle.
- standard math The stochastic interpolant flow framework is a valid generative formulation.
invented entities (3)
-
Proxel representation (multi-channel voxel density)
-
Chain flow vector field
-
ProxCLR embedding model
Cite this review
Pith. "Pith review of ProxelGen: Generating Proteins as 3D Densities." pith.science (2026). https://pith.science/paper/XQGBFMO7
@misc{pith2026250619820,
author = {Pith},
title = {Pith review of: ProxelGen: Generating Proteins as 3D Densities},
year = {2026},
howpublished = {\url{https://pith.science/paper/XQGBFMO7}},
note = {Machine review of arXiv:2506.19820}
}
read the original abstract
We develop ProxelGen, a protein structure generative model that operates on 3D densities as opposed to the prevailing 3D point cloud representations. Representing proteins as voxelized densities, or proxels, enables new tasks and conditioning capabilities. We generate proteins encoded as proxels via a 3D CNN-based VAE in conjunction with a diffusion model operating on its latent space. Compared to state-of-the-art models, ProxelGen's samples achieve higher novelty, better FID scores, and the same level of designability as the training set. ProxelGen's advantages are demonstrated in a standard motif scaffolding benchmark, and we show how 3D density-based generation allows for more flexible shape conditioning.
Figures
Figures from the paper (5 more)
Forward citations
Cited by 3 Pith papers
-
La-Proteina: Atomistic Protein Generation via Partially Latent Flow Matching
La-Proteina generates full-atom protein structures and sequences via flow matching over an explicit alpha-carbon backbone plus fixed-size per-residue latents, achieving state-of-the-art co-designability and scaling to...
-
Design-CP: Context Parallelism for Design of Protein Nanoparticles
Context-parallel inference for RFdiffusion 3 enables end-to-end all-atom design of large symmetric protein nanoparticles on multi-GPU hardware without retraining.
-
InertialAR: Autoregressive 3D Molecule Generation with Inertial Frames
An autoregressive transformer with inertial-frame tokenization and geometric rotary positional encoding reports state-of-the-art validity and stability on QM9, GEOM-Drugs, and B3LYP, plus strong functional-group-condi...
Reference graph
Works this paper leans on
-
[1]
Stochastic interpolants: A unifying framework for flows and diffusions
Michael S Albergo, Nicholas M Boffi, and Eric Vanden-Eijnden. Stochastic interpolants: A unifying framework for flows and diffusions. arXiv preprint arXiv:2303.08797, 2023
arXiv 2023
-
[2]
Building normalizing flows with stochastic interpolants
Michael Samuel Albergo and Eric Vanden-Eijnden. Building normalizing flows with stochastic interpolants. In The Eleventh International Conference on Learning Representations, 2022
work page 2022
-
[3]
Inigo Barrio-Hernandez, Jingi Yeo, J \"u rgen J \"a nes, Milot Mirdita, Cameron L. M. Gilchrist, Tanita Wein, Mihaly Varadi, Sameer Velankar, Pedro Beltrao, and Martin Steinegger. Clustering predicted structures at the scale of the known protein universe. Nature, 622 0 (7983): 0 637--645, Oct 2023. ISSN 1476-4687. doi:10.1038/s41586-023-06510-w. URL https...
-
[4]
Se(3)-stochastic flow matching for protein backbone generation, 2024
Avishek Joey Bose, Tara Akhound-Sadegh, Guillaume Huguet, Kilian Fatras, Jarrid Rector-Brooks, Cheng-Hao Liu, Andrei Cristian Nica, Maksym Korablyov, Michael Bronstein, and Alexander Tong. Se(3)-stochastic flow matching for protein backbone generation, 2024. URL https://arxiv.org/abs/2310.02391
arXiv 2024
-
[5]
A simple framework for contrastive learning of visual representations
Ting Chen, Simon Kornblith, Mohammad Norouzi, and Geoffrey Hinton. A simple framework for contrastive learning of visual representations. In International conference on machine learning, pages 1597--1607. PmLR, 2020
2020
-
[6]
Chu, Lucy Cheng, Gina El Nesr, Minkai Xu, and Po-Ssu Huang
Alexander E. Chu, Lucy Cheng, Gina El Nesr, Minkai Xu, and Po-Ssu Huang. An all-atom protein generative model. bioRxiv, 2023. doi:10.1101/2023.05.24.542194. URL https://www.biorxiv.org/content/early/2023/05/25/2023.05.24.542194
-
[7]
E(3)-equivariant models cannot learn chirality: Field-based molecular generation, 2025
Alexandru Dumitrescu, Dani Korpela, Markus Heinonen, Yogesh Verma, Valerii Iakovlev, Vikas Garg, and Harri Lähdesmäki. E(3)-equivariant models cannot learn chirality: Field-based molecular generation, 2025. URL https://arxiv.org/abs/2402.15864
arXiv 2025
-
[8]
Protein fid: Improved evaluation of protein structure generative models, 2025
Felix Faltings, Hannes Stark, Tommi Jaakkola, and Regina Barzilay. Protein fid: Improved evaluation of protein structure generative models, 2025. URL https://arxiv.org/abs/2505.08041
arXiv 2025
Show all 34 references
-
[9]
A latent diffusion model for protein structure generation, 2023
Cong Fu, Keqiang Yan, Limei Wang, Wing Yee Au, Michael McThrow, Tao Komikado, Koji Maruhashi, Kanji Uchino, Xiaoning Qian, and Shuiwang Ji. A latent diffusion model for protein structure generation, 2023. URL https://arxiv.org/abs/2305.04120
2023 arXiv
-
[10]
Proteina: Scaling flow-based protein structure generative models, 2025
Tomas Geffner, Kieran Didi, Zuobai Zhang, Danny Reidenbach, Zhonglin Cao, Jason Yim, Mario Geiger, Christian Dallago, Emine Kucukbenli, Arash Vahdat, and Karsten Kreis. Proteina: Scaling flow-based protein structure generative models, 2025. URL https://arxiv.org/abs/2503.00710
2025 arXiv
-
[11]
Gans trained by a two time-scale update rule converge to a local nash equilibrium
Martin Heusel, Hubert Ramsauer, Thomas Unterthiner, Bernhard Nessler, and Sepp Hochreiter. Gans trained by a two time-scale update rule converge to a local nash equilibrium. Advances in neural information processing systems, 30, 2017
2017
-
[12]
Auto-encoding variational bayes, 2022
Diederik P Kingma and Max Welling. Auto-encoding variational bayes, 2022. URL https://arxiv.org/abs/1312.6114
2022 arXiv
-
[13]
Generating novel, designable, and diverse protein structures by equivariantly diffusing oriented residue clouds, 2023
Yeqing Lin and Mohammed AlQuraishi. Generating novel, designable, and diverse protein structures by equivariantly diffusing oriented residue clouds, 2023. URL https://arxiv.org/abs/2301.12485
2023 arXiv
-
[14]
Out of many, one: Designing and scaffolding proteins at the scale of the structural universe with genie 2, 2024
Yeqing Lin, Minji Lee, Zhao Zhang, and Mohammed AlQuraishi. Out of many, one: Designing and scaffolding proteins at the scale of the structural universe with genie 2, 2024. URL https://arxiv.org/abs/2405.15489
2024 arXiv
-
[15]
Flow matching for generative modeling
Yaron Lipman, Ricky TQ Chen, Heli Ben-Hamu, Maximilian Nickel, and Matt Le. Flow matching for generative modeling. arXiv preprint arXiv:2210.02747, 2022
2022 arXiv
-
[16]
Flow straight and fast: Learning to generate and transfer data with rectified flow
Xingchao Liu, Chengyue Gong, and Qiang Liu. Flow straight and fast: Learning to generate and transfer data with rectified flow. arXiv preprint arXiv:2209.03003, 2022
2022 arXiv
-
[17]
Lu, Wilson Yan, Sarah A
Amy X. Lu, Wilson Yan, Sarah A. Robinson, Simon Kelow, Kevin K. Yang, Vladimir Gligorijevic, Kyunghyun Cho, Richard Bonneau, Pieter Abbeel, and Nathan C. Frey. All-atom protein generation with latent diffusion. bioRxiv, 2025 a . doi:10.1101/2024.12.02.626353. URL https://www.b...
2025 doi
-
[18]
Assessing generative model coverage of protein structures with shapes
Tianyu Lu, Melissa Liu, Yilin Chen, Jinho Kim, and Po-Ssu Huang. Assessing generative model coverage of protein structures with shapes. bioRxiv, pages 2025--01, 2025 b
2025
-
[19]
Bronstein, and Jinbo Xu
Matt McPartlon, C \'e line Marquet, Tomas Geffner, Daniel Kovtun, Alexander Goncearenco, Zachary Carpenter, Luca Naef, Michael M. Bronstein, and Jinbo Xu. Bridging sequence and structure: Latent diffusion for conditional protein generation, 2024. URL https://openreview.net/for...
2024
-
[20]
Pinheiro, Arian Jamasb, Omar Mahmood, Vishnu Sresht, and Saeed Saremi
Pedro O. Pinheiro, Arian Jamasb, Omar Mahmood, Vishnu Sresht, and Saeed Saremi. Structure-based drug design by denoising voxel grids, 2024 a . URL https://arxiv.org/abs/2405.03961
2024 arXiv
-
[21]
Pinheiro, Joshua Rackers, Joseph Kleinhenz, Michael Maser, Omar Mahmood, Andrew Martin Watkins, Stephen Ra, Vishnu Sresht, and Saeed Saremi
Pedro O. Pinheiro, Joshua Rackers, Joseph Kleinhenz, Michael Maser, Omar Mahmood, Andrew Martin Watkins, Stephen Ra, Vishnu Sresht, and Saeed Saremi. 3d molecule generation by denoising voxel grids, 2024 b . URL https://arxiv.org/abs/2306.07473
2024 arXiv
-
[22]
High-resolution image synthesis with latent diffusion models, 2022
Robin Rombach, Andreas Blattmann, Dominik Lorenz, Patrick Esser, and Björn Ommer. High-resolution image synthesis with latent diffusion models, 2022. URL https://arxiv.org/abs/2112.10752
2022 arXiv
-
[23]
Score-based generative modeling through stochastic differential equations
Yang Song, Jascha Sohl-Dickstein, Diederik P Kingma, Abhishek Kumar, Stefano Ermon, and Ben Poole. Score-based generative modeling through stochastic differential equations. arXiv preprint arXiv:2011.13456, 2020
2011 arXiv
-
[24]
Score-based generative modeling in latent space, 2021
Arash Vahdat, Karsten Kreis, and Jan Kautz. Score-based generative modeling in latent space, 2021. URL https://arxiv.org/abs/2106.05931
2021 arXiv
-
[25]
Alphafold protein structure database in 2024: providing structure coverage for over 214 million protein sequences
Mihaly Varadi, Damian Bertoni, Paulyna Magana, Urmila Paramval, Ivanna Pidruchna, Malarvizhi Radhakrishnan, Maxim Tsenkov, Sreenath Nair, Milot Mirdita, Jingi Yeo, Oleg Kovalevskiy, Kathryn Tunyasuvunakool, Agata Laydon, Augustin Žídek, Hamish Tomlinson, Dhavanthi Hariharan, J...
2024
-
[26]
Generating highly designable proteins with geometric algebra flow matching, 2024
Simon Wagner, Leif Seute, Vsevolod Viliuga, Nicolas Wolf, Frauke Gräter, and Jan Stühmer. Generating highly designable proteins with geometric algebra flow matching, 2024. URL https://arxiv.org/abs/2411.05238
2024 arXiv
-
[27]
Watson, David Juergens, Nathaniel R
Joseph L. Watson, David Juergens, Nathaniel R. Bennett, Brian L. Trippe, Jason Yim, Helen E. Eisenach, Woody Ahern, Andrew J. Borst, Robert J. Ragotte, Lukas F. Milles, Basile I. M. Wicky, Nikita Hanikel, Samuel J. Pellock, Alexis Courbet, William Sheffler, Jue Wang, Preetham ...
2023
-
[28]
Jason Yim, Andrew Campbell, Andrew Y. K. Foong, Michael Gastegger, José Jiménez-Luna, Sarah Lewis, Victor Garcia Satorras, Bastiaan S. Veeling, Regina Barzilay, Tommi Jaakkola, and Frank Noé. Fast protein backbone generation with se(3) flow matching, 2023 a . URL https://arxiv...
2023 arXiv
-
[29]
Trippe, Valentin De Bortoli, Emile Mathieu, Arnaud Doucet, Regina Barzilay, and Tommi Jaakkola
Jason Yim, Brian L. Trippe, Valentin De Bortoli, Emile Mathieu, Arnaud Doucet, Regina Barzilay, and Tommi Jaakkola. Se(3) diffusion model with application to protein backbone generation, 2023 b . URL https://arxiv.org/abs/2302.02277
2023 arXiv
-
[30]
Jason Yim, Andrew Campbell, Emile Mathieu, Andrew Y. K. Foong, Michael Gastegger, José Jiménez-Luna, Sarah Lewis, Victor Garcia Satorras, Bastiaan S. Veeling, Frank Noé, Regina Barzilay, and Tommi S. Jaakkola. Improved motif-scaffolding with se(3) flow matching, 2024 a . URL h...
2024 arXiv
-
[31]
Jaakkola
Jason Yim, Hannes Stärk, Gabriele Corso, Bowen Jing, Regina Barzilay, and Tommi S. Jaakkola. Diffusion models in protein structure and docking. WIREs Computational Molecular Science, 14 0 (2): 0 e1711, 2024 b . doi:https://doi.org/10.1002/wcms.1711. URL https://wires.onlinelib...
2024 doi
-
[32]
Hierarchical protein backbone generation with latent and structure diffusion, 2025
Jason Yim, Marouane Jaakik, Ge Liu, Jacob Gershon, Karsten Kreis, David Baker, Regina Barzilay, and Tommi Jaakkola. Hierarchical protein backbone generation with latent and structure diffusion, 2025. URL https://arxiv.org/abs/2504.09374
2025 arXiv
-
[33]
Ecloudgen: Leveraging electron clouds as a latent variable to scale up structure-based molecular design
Odin Zhang, Jieyu Jin, Zhenxing Wu, Jintu Zhang, Po Yuan, Haitao Lin, Haiyang Zhong, Xujun Zhang, Chenqing Hua, Weibo Zhao, Zhengshuo Zhang, Kejun Ying, Yufei Huang, Huifeng Zhao, Yuntao Yu, Yu Kang, Peichen Pan, Jike Wang, Dong Guo, Shuangjia Zheng, Chang-Yu Hsieh, and Tingju...
2024 doi
-
[34]
Tm-align: a protein structure alignment algorithm based on the tm-score
Yang Zhang and Jeffrey Skolnick. Tm-align: a protein structure alignment algorithm based on the tm-score. Nucleic Acids Research, 33 0 (7): 0 2302--2309, 01 2005. ISSN 0305-1048. doi:10.1093/nar/gki524. URL https://doi.org/10.1093/nar/gki524
2005 doi
Reviewed August 15, 2026 · model on record in the stance chip above.
Discussion (0). Continue with ORCID to comment.