2 citations · 4 across the 5 of their papers we have counts for
3 papers · 1 filter
Nemotron 3 Ultra: Open, Efficient Mixture-of-Experts Hybrid Mamba-Transformer Model for Agentic Reasoning
NVIDIA, :, Aaron Blakeman +571
We introduce Nemotron 3 Ultra, a 550 billion total and 55 billion active parameter Mixture-of-Experts Hybrid Mamba-Attention language model. We pre-trained Nemotron 3 Ultra on 20 t…
Probabilistic Contrastive Pretraining for Multi-task ADME Property Prediction
Yifan Xue, Srimukh Prasad Veccham, Saee Paliwal +2
Accurate prediction of absorption, distribution, metabolism, and excretion (ADME) properties is critical to drug discovery, but remains challenging because ADME endpoints are noisy…
Exploring Synthesizable Chemical Space with Iterative Pathway Refinements
Seul Lee, Karsten Kreis, Srimukh Prasad Veccham +5
A well-known pitfall of molecular generative models is that they are not guaranteed to generate synthesizable molecules. Existing solutions for this problem often struggle to effec…