default search action
Somshubra Majumdar
Person information
Refine list
refinements active!
zoomed in on ?? of ?? records
view refined list in
export refined list as
2020 – today
- 2024
- [c14]Vahid Noroozi, Somshubra Majumdar, Ankur Kumar, Jagadeesh Balam, Boris Ginsburg:
Stateful Conformer with Cache-Based Inference for Streaming Automatic Speech Recognition. ICASSP 2024: 12041-12045 - [c13]Nithin Rao Koluguri, Samuel Kriman, Georgy Zelenfroind, Somshubra Majumdar, Dima Rekesh, Vahid Noroozi, Jagadeesh Balam, Boris Ginsburg:
Investigating End-to-End ASR Architectures for Long Form Audio Transcription. ICASSP 2024: 13366-13370 - [i21]Bo Adler, Niket Agarwal, Ashwath Aithal, Dong H. Anh, Pallab Bhattacharya, Annika Brundyn, Jared Casper, Bryan Catanzaro, Sharon Clay, Jonathan M. Cohen, Sirshak Das, Ayush Dattagupta, Olivier Delalleau, Leon Derczynski, Yi Dong, Daniel Egert, Ellie Evans, Aleksander Ficek, Denys Fridman, Shaona Ghosh, Boris Ginsburg, Igor Gitman, Tomasz Grzegorzek, Robert Hero, Jining Huang, Vibhu Jawa, Joseph Jennings, Aastha Jhunjhunwala, John Kamalu, Sadaf Khan, Oleksii Kuchaiev, Patrick LeGresley, Hui Li, Jiwei Liu, Zihan Liu, Eileen Long, Ameya Sunil Mahabaleshwarkar, Somshubra Majumdar, James Maki, Miguel Martinez, Maer Rodrigues de Melo, Ivan Moshkov, Deepak Narayanan, Sean Narenthiran, Jesus Navarro, Phong Nguyen, Osvald Nitski, Vahid Noroozi, Guruprasad Nutheti, Christopher Parisien, Jupinder Parmar, Mostofa Patwary, Krzysztof Pawelec, Wei Ping, Shrimai Prabhumoye, Rajarshi Roy, Trisha Saar, Vasanth Rao Naik Sabavat, Sanjeev Satheesh, Jane Polak Scowcroft, Jason Sewall, Pavel Shamis, Gerald Shen, Mohammad Shoeybi, Dave Sizer, Misha Smelyanskiy, Felipe Soares, Makesh Narsimhan Sreedhar, Dan Su, Sandeep Subramanian, Shengyang Sun, Shubham Toshniwal, Hao Wang, Zhilin Wang, Jiaxuan You, Jiaqi Zeng, Jimmy Zhang, Jing Zhang, Vivienne Zhang, Yian Zhang, Chen Zhu:
Nemotron-4 340B Technical Report. CoRR abs/2406.11704 (2024) - [i20]Vahid Noroozi, Zhehuai Chen, Somshubra Majumdar, Steve Huang, Jagadeesh Balam, Boris Ginsburg:
Instruction Data Generation and Unsupervised Adaptation for Speech Language Models. CoRR abs/2406.12946 (2024) - [i19]Krishna C. Puvvada, Piotr Zelasko, He Huang, Oleksii Hrinchuk, Nithin Rao Koluguri, Kunal Dhawan, Somshubra Majumdar, Elena Rastorgueva, Zhehuai Chen, Vitaly Lavrukhin, Jagadeesh Balam, Boris Ginsburg:
Less is More: Accurate Speech Recognition & Translation without Web-Scale Data. CoRR abs/2406.19674 (2024) - [i18]Somshubra Majumdar, Vahid Noroozi, Sean Narenthiran, Aleksander Ficek, Jagadeesh Balam, Boris Ginsburg:
Genetic Instruct: Scaling up Synthetic Generation of Coding Instructions for Large Language Models. CoRR abs/2407.21077 (2024) - [i17]Weiqing Wang, Kunal Dhawan, Taejin Park, Krishna C. Puvvada, Ivan Medennikov, Somshubra Majumdar, He Huang, Jagadeesh Balam, Boris Ginsburg:
Resource-Efficient Adaptation of Speech Foundation Models for Multi-Speaker ASR. CoRR abs/2409.01438 (2024) - 2023
- [c12]Dima Rekesh, Nithin Rao Koluguri, Samuel Kriman, Somshubra Majumdar, Vahid Noroozi, He Huang, Oleksii Hrinchuk, Krishna C. Puvvada, Ankur Kumar, Jagadeesh Balam, Boris Ginsburg:
Fast Conformer With Linearly Scalable Attention For Efficient Speech Recognition. ASRU 2023: 1-8 - [c11]Hainan Xu, Fei Jia, Somshubra Majumdar, Shinji Watanabe, Boris Ginsburg:
Multi-Blank Transducers for Speech Recognition. ICASSP 2023: 1-5 - [c10]Hainan Xu, Fei Jia, Somshubra Majumdar, He Huang, Shinji Watanabe, Boris Ginsburg:
Efficient Sequence Transduction by Jointly Predicting Tokens and Durations. ICML 2023: 38462-38484 - [i16]Hainan Xu, Fei Jia, Somshubra Majumdar, He Huang, Shinji Watanabe, Boris Ginsburg:
Efficient Sequence Transduction by Jointly Predicting Tokens and Durations. CoRR abs/2304.06795 (2023) - [i15]Dima Rekesh, Samuel Kriman, Somshubra Majumdar, Vahid Noroozi, He Huang, Oleksii Hrinchuk, Ankur Kumar, Boris Ginsburg:
Fast Conformer with Linearly Scalable Attention for Efficient Speech Recognition. CoRR abs/2305.05084 (2023) - [i14]Nithin Rao Koluguri, Samuel Kriman, Georgy Zelenfroind, Somshubra Majumdar, Dima Rekesh, Vahid Noroozi, Jagadeesh Balam, Boris Ginsburg:
Investigating End-to-End ASR Architectures for Long Form Audio Transcription. CoRR abs/2309.09950 (2023) - [i13]Vahid Noroozi, Somshubra Majumdar, Ankur Kumar, Jagadeesh Balam, Boris Ginsburg:
Stateful Conformer with Cache-based Inference for Streaming Automatic Speech Recognition. CoRR abs/2312.17279 (2023) - 2022
- [c9]Aleksandr Laptev, Somshubra Majumdar, Boris Ginsburg:
CTC Variations Through New WFST Topologies. INTERSPEECH 2022: 1041-1045 - [c8]Oleksii Hrinchuk, Vahid Noroozi, Ashwinkumar Ganesan, Sarah Campbell, Sandeep Subramanian, Somshubra Majumdar, Oleksii Kuchaiev:
NVIDIA NeMo Offline Speech Translation Systems for IWSLT 2022. IWSLT@ACL 2022: 225-231 - [c7]Somshubra Majumdar, Shantanu Acharya, Vitaly Lavrukhin, Boris Ginsburg:
Damage Control During Domain Adaptation for Transducer Based Automatic Speech Recognition. SLT 2022: 130-135 - [i12]Somshubra Majumdar, Shantanu Acharya, Vitaly Lavrukhin, Boris Ginsburg:
Damage Control During Domain Adaptation for Transducer Based Automatic Speech Recognition. CoRR abs/2210.03255 (2022) - [i11]Hainan Xu, Fei Jia, Somshubra Majumdar, Shinji Watanabe, Boris Ginsburg:
Multi-blank Transducers for Speech Recognition. CoRR abs/2211.03541 (2022) - 2021
- [j5]Fazle Karim, Somshubra Majumdar, Houshang Darabi:
Adversarial Attacks on Time Series. IEEE Trans. Pattern Anal. Mach. Intell. 43(10): 3309-3320 (2021) - [c6]Fei Jia, Somshubra Majumdar, Boris Ginsburg:
MarbleNet: Deep 1D Time-Channel Separable Convolutional Neural Network for Voice Activity Detection. ICASSP 2021: 6818-6822 - [c5]Patrick K. O'Neill, Vitaly Lavrukhin, Somshubra Majumdar, Vahid Noroozi, Yuekai Zhang, Oleksii Kuchaiev, Jagadeesh Balam, Yuliya Dovzhenko, Keenan Freyberg, Michael D. Shulman, Boris Ginsburg, Shinji Watanabe, Georg Kucsko:
SPGISpeech: 5, 000 Hours of Transcribed Financial Audio for Fully Formatted End-to-End Speech Recognition. Interspeech 2021: 1434-1438 - [i10]Patrick K. O'Neill, Vitaly Lavrukhin, Somshubra Majumdar, Vahid Noroozi, Yuekai Zhang, Oleksii Kuchaiev, Jagadeesh Balam, Yuliya Dovzhenko, Keenan Freyberg, Michael D. Shulman, Boris Ginsburg, Shinji Watanabe, Georg Kucsko:
SPGISpeech: 5, 000 hours of transcribed financial audio for fully formatted end-to-end speech recognition. CoRR abs/2104.02014 (2021) - [i9]Aleksei Kalinov, Somshubra Majumdar, Jagadeesh Balam, Boris Ginsburg:
CarneliNet: Neural Mixture Model for Automatic Speech Recognition. CoRR abs/2107.10708 (2021) - [i8]Aleksandr Laptev, Somshubra Majumdar, Boris Ginsburg:
CTC Variations Through New WFST Topologies. CoRR abs/2110.03098 (2021) - 2020
- [c4]Somshubra Majumdar, Boris Ginsburg:
MatchboxNet: 1D Time-Channel Separable Convolutional Neural Network Architecture for Speech Commands Recognition. INTERSPEECH 2020: 3356-3360 - [i7]Fei Jia, Somshubra Majumdar, Boris Ginsburg:
MarbleNet: Deep 1D Time-Channel Separable Convolutional Neural Network for Voice Activity Detection. CoRR abs/2010.13886 (2020)
2010 – 2019
- 2019
- [j4]Fazle Karim, Somshubra Majumdar, Houshang Darabi:
Insights Into LSTM Fully Convolutional Networks for Time Series Classification. IEEE Access 7: 67718-67725 (2019) - [j3]Fazle Karim, Somshubra Majumdar, Houshang Darabi, Samuel Harford:
Multivariate LSTM-FCNs for time series classification. Neural Networks 116: 237-245 (2019) - [i6]Fazle Karim, Somshubra Majumdar, Houshang Darabi:
Adversarial Attacks on Time Series. CoRR abs/1902.10755 (2019) - [i5]Fazle Karim, Somshubra Majumdar, Houshang Darabi:
Insights into LSTM Fully Convolutional Networks for Time Series Classification. CoRR abs/1902.10756 (2019) - 2018
- [j2]Fazle Karim, Somshubra Majumdar, Houshang Darabi, Shun Chen:
LSTM Fully Convolutional Networks for Time Series Classification. IEEE Access 6: 1662-1669 (2018) - [j1]Piotr Chudzik, Somshubra Majumdar, Francesco Calivá, Bashir Al-Diri, Andrew Hunter:
Microaneurysm detection using fully convolutional neural networks. Comput. Methods Programs Biomed. 158: 185-192 (2018) - [c3]Maryam Pishgar, Fazle Karim, Somshubra Majumdar, Houshang Darabi:
Pathological Voice Classification Using Mel-Cepstrum Vectors and Support Vector Machine. IEEE BigData 2018: 5267-5271 - [c2]Piotr Chudzik, Somshubra Majumdar, Francesco Calivá, Bashir Al-Diri, Andrew Hunter:
Microaneurysm detection using deep learning and interleaved freezing. Medical Imaging: Image Processing 2018: 105741I - [c1]Piotr Chudzik, Somshubra Majumdar, Francesco Calivá, Bashir Al-Diri, Andrew Hunter:
Exudate segmentation using fully convolutional neural networks and inception modules. Medical Imaging: Image Processing 2018: 1057430 - [i4]Fazle Karim, Somshubra Majumdar, Houshang Darabi, Samuel Harford:
Multivariate LSTM-FCNs for Time Series Classification. CoRR abs/1801.04503 (2018) - [i3]Somshubra Majumdar, Amlaan Bhoi, Ganesh Jagadeesan:
A Comprehensive Comparison between Neural Style Transfer and Universal Style Transfer. CoRR abs/1806.00868 (2018) - [i2]Maryam Pishgar, Fazle Karim, Somshubra Majumdar, Houshang Darabi:
Pathological Voice Classification Using Mel-Cepstrum Vectors and Support Vector Machine. CoRR abs/1812.07729 (2018) - 2017
- [i1]Fazle Karim, Somshubra Majumdar, Houshang Darabi, Shun Chen:
LSTM Fully Convolutional Networks for Time Series Classification. CoRR abs/1709.05206 (2017)
Coauthor Index
manage site settings
To protect your privacy, all features that rely on external API calls from your browser are turned off by default. You need to opt-in for them to become active. All settings here will be stored as cookies with your web browser. For more information see our F.A.Q.
Unpaywalled article links
Add open access links from to the list of external document links (if available).
Privacy notice: By enabling the option above, your browser will contact the API of unpaywall.org to load hyperlinks to open access articles. Although we do not have any reason to believe that your call will be tracked, we do not have any control over how the remote server uses your data. So please proceed with care and consider checking the Unpaywall privacy policy.
Archived links via Wayback Machine
For web page which are no longer available, try to retrieve content from the of the Internet Archive (if available).
Privacy notice: By enabling the option above, your browser will contact the API of archive.org to check for archived content of web pages that are no longer available. Although we do not have any reason to believe that your call will be tracked, we do not have any control over how the remote server uses your data. So please proceed with care and consider checking the Internet Archive privacy policy.
Reference lists
Add a list of references from , , and to record detail pages.
load references from crossref.org and opencitations.net
Privacy notice: By enabling the option above, your browser will contact the APIs of crossref.org, opencitations.net, and semanticscholar.org to load article reference information. Although we do not have any reason to believe that your call will be tracked, we do not have any control over how the remote server uses your data. So please proceed with care and consider checking the Crossref privacy policy and the OpenCitations privacy policy, as well as the AI2 Privacy Policy covering Semantic Scholar.
Citation data
Add a list of citing articles from and to record detail pages.
load citations from opencitations.net
Privacy notice: By enabling the option above, your browser will contact the API of opencitations.net and semanticscholar.org to load citation information. Although we do not have any reason to believe that your call will be tracked, we do not have any control over how the remote server uses your data. So please proceed with care and consider checking the OpenCitations privacy policy as well as the AI2 Privacy Policy covering Semantic Scholar.
OpenAlex data
Load additional information about publications from .
Privacy notice: By enabling the option above, your browser will contact the API of openalex.org to load additional information. Although we do not have any reason to believe that your call will be tracked, we do not have any control over how the remote server uses your data. So please proceed with care and consider checking the information given by OpenAlex.
last updated on 2024-10-22 20:15 CEST by the dblp team
all metadata released as open data under CC0 1.0 license
see also: Terms of Use | Privacy Policy | Imprint