# Year Type Publication
[1] 2026 Journal
Last-Iterate Convergence of General Parameterized Policies in Constrained MDPs
Transactions on Machine Learning Research
[2] 2025 Conference
A Sharper Global Convergence Analysis for Average Reward Reinforcement Learning via an Actor-Critic Approach
International Conference on Machine Learning (ICML)
[3] 2025 Conference
Finite-Sample Analysis of Policy Evaluation for Robust Average Reward Reinforcement Learning
Neural Information Processing Systems (NeurIPS)
[4] 2025 Conference
Global Convergence for Average Reward Constrained MDPs with Primal-Dual Actor Critic Algorithm
Neural Information Processing Systems (NeurIPS)
[5] 2025 Journal
MOORL: A Framework for Integrating Offline-Online Reinforcement Learning
Transactions on Machine Learning Research
[6] 2025 Conference
Order-Optimal Global Convergence for Average Reward Actor-Critic with General Policy and Neural Critic Parametrization
Conference on Uncertainty in Artificial Intelligence (UAI)
[7] 2025 Conference
Order-Optimal Regret with Novel Policy Gradient Approaches in Infinite Horizon Average Reward MDPs
International Conference on Artificial Intelligence and Statistics (AISTATS)
[8] 2025 Conference
Uplink NOMA-Aided Multi-Device Multi-Target Integrated Sensing and Communication
National Conference on Communications (NCC)
[9] 2024 Book
Constrained Reinforcement Learning with Average Reward Objective: Model-Based and Model-Free Algorithms
Foundations and TrendsĀ® in Optimization (Now Publishers)
[10] 2024 Journal
Cooperating Graph Neural Networks with Deep Reinforcement Learning for Vaccine Prioritization
IEEE Journal of Biomedical and Health Informatics
[11] 2024 Conference
Improved Sample Complexity Analysis of Natural Policy Gradient Algorithm with General Parameterization for Infinite Horizon Discounted Reward Markov Decision Processes
International Conference on Artificial Intelligence and Statistics (AISTATS)
[12] 2024 Conference
Learning General Parameterized Policies for Infinite Horizon Average Reward Constrained MDPs via Primal-Dual Policy Gradient Algorithm
Neural Information Processing Systems (NeurIPS)
[13] 2024 Journal
Mean-Field Approximation of Cooperative Constrained Multi-Agent Reinforcement Learning (CMARL)
Journal of Machine Learning Research
[14] 2024 Journal
Near-Perfect Coverage Manifold Estimation in Cellular Networks via Conditional GAN
IEEE Networking Letters
[15] 2024 Conference
Regret Analysis of Policy Gradient Algorithm for Infinite Horizon Average Reward Markov Decision Processes
AAAI Conference on Artificial Intelligence
[16] 2024 Conference
Sample-Efficient Constrained Reinforcement Learning with General Parameterization
Neural Information Processing Systems (NeurIPS)
[17] 2024 Journal
Terrain-Based Coverage Manifold Estimation: Machine Learning, Stochastic Geometry, or Simulation?
IEEE Open Journal of the Communications Society
[18] 2023 Journal
Coalition Formation for Outsourced Spectrum Sensing in Cognitive Radio Network
IEEE Transactions on Cognitive Communications and Networking
[19] 2023 Journal
Mean-Field Control based Approximation of Multi-Agent Reinforcement Learning in Presence of a Non-decomposable Shared Global State
Transactions on Machine Learning Research
[20] 2023 Journal
Reinforcement Learning with Delayed, Composite, and Partially Anonymous Reward
Transactions on Machine Learning Research
[21] 2022 Conference
Can Mean Field Control (MFC) Approximate Cooperative Multi Agent Reinforcement Learning (MARL) with Non-Uniform Interaction?
Conference on Uncertainty in Artificial Intelligence (UAI)
[22] 2022 Journal
Deep Learning based Coverage and Rate Manifold Estimation in Cellular Networks
IEEE Transactions on Cognitive Communications and Networking
[23] 2022 Journal
On the Approximation of Cooperative Heterogeneous Multi-Agent Reinforcement Learning (MARL) using Mean Field Control (MFC)
Journal of Machine Learning Research
[24] 2022 Journal
On the Near-Optimality of Local Policies in Large Cooperative Multi-Agent Reinforcement Learning
Transactions on Machine Learning Research
[25] 2022 Journal
Queuing Analysis of QoS Aware Microwave Power Transfer Enabled CR-IoT Network
IEEE Transactions on Wireless Communications
[26] 2022 Journal
Sustainability Analysis of Opportunistic CR-IoT Network Employing Microwave Power Transfer
IEEE Transactions on Cognitive Communications and Networking
[27] 2021 Journal
Economic Analysis of Cognitive Underlay Networks: A Nash Bargaining Based Approach
IEEE Transactions on Vehicular Technology
[28] 2021 Journal
On Exact Distribution of Poisson-Voronoi Area in K-Tier HetNets With Generalized Association Rule
IEEE Communications Letters
[29] 2021 Conference
On the Approximation of Cooperative Heterogeneous Multi-Agent Reinforcement Learning (MARL) using Mean Field Control (MFC)
Cooperative AI workshop at Neural Information Processing Systems (NeurIPS)
[30] 2020 Journal
Economics of Resilient TWDM PONs
IEEE/OSA Journal of Lightwave Technology
[31] 2019 Journal
Economic analysis of TWDM PONs: A sustainability and policy-making perspective
IEEE/OSA Journal of Optical Communications and Networking
[32] 2019 Journal
Economics of TWDM PONs with Nonlinear Pricing
IEEE Communications Letters
[33] 2019 Conference
Nash Bargaining Based Economic Analysis of Cognitive Cellular Networks
IEEE International Conference on Communications (ICC)
[34] 2019 Journal
Nash Bargaining Based Economic Analysis of Opportunistic Cognitive Cellular Networks
IEEE Transactions on Cognitive Communications and Networking
[35] 2019 Conference
Predation Blocking Strategies in Real Cellular Networks and its impact on Spectrum Revenue
IEEE Vehicular Technology Conference (VTC)
[36] 2017 Journal
Blocking predation in cellular monopoly through non-linear spectrum pricing
IEEE Communications Letters
[37] 2017 Journal
Traffic-Aware Green Cognitive Radio
Physical Communication
[38] 2017 Journal
Uplink User Process in Poisson Cellular Network
IEEE Communications Letters