Technology
59b1deff341edb0b76ace57820cef237-AuthorFeedback.pdf
Indeed, the results in Table 1, which shows13 the mean absolute percentage errors (MAPE), demonstrates this. The ac-14 curacy of neural ODE for the Poisson process is on par with our neural15 JSDE. However, for the Hawkes process (Exponential), Hawkes process16 (Power-Law), and self-correcting process, neural ODE gives much larger17 predictions errors. Forthesocial/medicaldatasets,weuseda20/64-24 dimensional latent state and parameterized the functions with two-hidden-layer MLPs with 32/64 hidden units. The time series modeling software that we used is designed for long event sequences and ignores the idle time after31 thelastevent.
daf8364f0715a41a469c677c0adc4754-Supplemental-Conference.pdf
Since weak learners perform only marginallybetter than random guesses, such subroutines constitute aweakerassumption than the availability of an accurate supervised learning oracle. Weprovethat the sample complexity and running time bounds of the proposed method do not explicitly dependonthenumberofstates. While existing results on boosting operate on convex losses, the value function over policies is non-convex.