An overview of Ilov3Splat. Left: Our method learns language-aligned and instance-aware features for 3D Gaussians, computed via compact multi-resolution hash encoding and lightweight projection MLPs. Right: Feature learning is guided by multi-view 2D signals, leveraging CLIP for language alignment, DINO for object boundary regularization, and SAM for instance-aware contrastive learning.