← Search

Florian Krach

2 accepted papers

2025

Filtered not Mixed: Filtering-Based Online Gating for Mixture of Large Language Models

ICLR 2025poster

We propose MoE-F — a formalized mechanism for combining N pre-trained expert Large Language Models (LLMs) in online time-series prediction tasks by adaptively forecasting the best weighting of LLM predictions at every time step. Our mechanism leverages the conditional information in each expert's ru…

2021

Neural Jump Ordinary Differential Equations: Consistent Continuous-Time Prediction and Filtering

ICLR 2021poster

Combinations of neural ODEs with recurrent neural networks (RNN), like GRU-ODE-Bayes or ODE-RNN are well suited to model irregularly observed time series. While those models outperform existing discrete-time approaches, no theoretical guarantees for their predictive capabilities are available. Assum…

Cited by 48SourcePDFScholar