SenseTime's co-founder and chief scientist Lin Dahua has indicated that a significant advancement in native multimodal artificial intelligence could be achieved within the next two years, according to KrASIA.
Multimodal AI refers to systems that can process and integrate multiple types of data inputs such as text, images, and audio, enhancing the capability and versatility of artificial intelligence applications.
This development is particularly relevant to Japanese markets, where AI-driven technologies are increasingly influencing sectors from financial services to automated trading, making breakthroughs in multimodal AI a potential catalyst for innovation across FX, crypto, and equities markets.
