← Search

Jun Jie Sim

2 accepted papers

2026

MOAI: Module-Optimizing Architecture for Non-Interactive Secure Transformer Inference

ICLR 2026poster

Privacy concerns have been raised in Large Language Models (LLM) inference when models are deployed in Cloud Service Providers (CSP). Homomorphic encryption (HE) offers a promising solution by enabling secure inference directly over encrypted inputs. However, the high computational overhead of HE re…

Cited by 0SourcecodeScholar
2026

Pisces: Cryptography-based Private Retrieval-Augmented Generation with Dual-Path Retrieval

ICLR 2026poster

Retrieval-augmented generation (RAG) enhances the response quality of large language models (LLMs) when handling domain-specific tasks, yet raises significant privacy concerns. This is because both the user query and documents within the knowledge base often contain sensitive or confidential informa…

Cited by 0SourcecodeScholar