Publications (13)
Towards Cooperation in Sequential Prisoner's Dilemmas: a Deep Multiagent Reinforcement Learning Approach
Weixun Wang, Jianye Hao, Yixi Wang +1
The Iterated Prisoner's Dilemma has guided research on social dilemmas for decades. However, it distinguishes between only two atomic actions: cooperate and defect. In real-world p…
Assessing AI vs Human-Authored Spear Phishing SMS Attacks: An Empirical Study
Jerson Francia, Derek Hansen, Ben Schooley +3
This paper explores the use of Large Language Models (LLMs) in spear phishing message generation and evaluates their performance compared to human-authored counterparts. Our pilot…
Learning to Shape Rewards using a Game of Two Partners
David Mguni, Taher Jafferjee, Jianhong Wang +9
Reward shaping (RS) is a powerful method in reinforcement learning (RL) for overcoming the problem of sparse or uninformative rewards. However, RS typically relies on manually engi…
Taming Multi-Agent Reinforcement Learning with Estimator Variance Reduction
Taher Jafferjee, Juliusz Ziomek, Tianpei Yang +6
Centralised training with decentralised execution (CT-DE) serves as the foundation of many leading multi-agent reinforcement learning (MARL) algorithms. Despite its popularity, it…
Thirty Meter Telescope International Observatory Detailed Science Case 2024
Warren Skidmore, Bob Kirshner, David Andersen +168
The Thirty Meter Telescope (TMT) International Observatory (TIO) will be a revolutionary leap forward in astronomical observing capabilities, enabling us to address some of the mos…
The Canada-France Ecliptic Plane Survey - Full Data Release: The orbital structure of the Kuiper belt
Jean-Marc Petit, J. John Kavelaars, Brett J. Gladman +14
We report the orbital distribution of the trans-neptunian objects (TNOs) discovered during the Canada-France Ecliptic Plane Survey, whose discovery phase ran from early 2003 until…
Electron intensity measurements by the Cluster/RAPID/IES instrument in Earths radiation belts and ring current
Artem Smirnov, Elena Kronberg, Flanck Latallerie +10
The Cluster mission, launched in 2000, has produced a large database of electron flux intensity measurements in the Earths magnetosphere by the Research with Adaptive Particle Imag…
Using PCA to Efficiently Represent State Spaces
William Curran, Tim Brys, Matthew Taylor +1
Reinforcement learning algorithms need to deal with the exponential growth of states and actions when exploring optimal control in high-dimensional spaces. This is known as the cur…
AVID: A Near-Major Post-Merger of Late-Type Dwarfs beneath a Regularly Rotating HI Disk (VCC 693)
Fujia Li, Hong-Xin Zhang, Elias Brinks +23
On the periphery of galaxy clusters, moderately high galaxy densities and velocity dispersions favour interactions and mergers that influence galaxy evolution prior to cluster infa…
Random matrix ensembles for -symmetric systems
Eva-Maria Graefe, Steve Mudute-Ndumbe, Matthew Taylor
Recently much effort has been made towards the introduction of non-Hermitian random matrix models respecting -symmetry. Here we show that there is a one-to-one correspondence b…
Decentralized Coordination of Distributed Energy Resources through Local Energy Markets and Deep Reinforcement Learning
Daniel May, Matthew Taylor, Petr Musilek
As distributed energy resources (DERs) grow, the electricity grid faces increased net load variability at the grid edge, impacting operability and reliability. Transactive energy,…
SpellBound: Defending Against Package Typosquatting
Matthew Taylor, Ruturaj K. Vaidya, Drew Davidson +2
Package managers for software repositories based on a single programming language are very common. Examples include npm (JavaScript), and PyPI (Python). These tools encourage code…
Observational Evidence for a Dark Side to NGC5128's Globular Cluster System
Matthew Taylor, Thomas Puzia, Matias Gomez +1
We present a study of the dynamical properties of 125 compact stellar systems (CSSs) in the nearby giant elliptical galaxy NGC5128, using high-resolution spectra (R 26,000) obtaine…