Google Gemini

Google Gemini is a family of advanced multimodal artificial intelligence (AI) models developed by Google. Introduced as the successor to Google Bard, Gemini is designed to be an AI assistant capable of understanding and processing various types of input, such as text, images, audio, video, and programming code. Since its official launch in February 2024, Gemini has evolved into one of the most comprehensive AI ecosystems, deeply integrated with Google services and available in various subscription tiers to meet the needs of users ranging from students to enterprises.

History and Development

The roots of Google Gemini lie in Google Bard, a chatbot based on the LaMDA language model that was launched in 2023. A major transformation took place in early 2024 when Bard was upgraded with a more advanced next-generation large language model (LLM). On February 8, 2024, Google officially announced the name change from Bard to Gemini, marking a strategic shift to unify all of its AI capabilities under a single brand umbrella. This change was also accompanied by the launch of the Gemini app for smartphones, which gradually replaced Google Assistant as the primary digital assistant on certain Android devices.

Throughout 2025 and 2026, Google continued to develop Gemini models at a high pace. The launch of models such as Gemini 3.5 Flash and Gemini 3.1 Flash-Lite demonstrated Google's commitment to providing models with a balance between high intelligence and low latency, especially for programming and agentic tasks. Recent innovations include the introduction of Gemini Omni Flash, a multimodal model capable of generating 720p-quality video from text descriptions or editing video through simple conversation. This model was released in public preview in late June 2026, marking a major step for Google in the field of generative content creation.

Core Capabilities and Features

Multimodality and Content Processing

Gemini's main feature is its capability as a multimodal model. Users can interact with Gemini through text, voice, and images. The Gemini app allows users to take photos and ask questions about objects around them. This capability is extended with the video-to-image feature on the latest model, in which Gemini can generate high-quality thumbnails or posters from uploaded video files. Gemini is also able to create Audio Overviews, turning documents and reports into engaging podcast-style discussions, making complex content easier to consume.

Integration and Extensions

Gemini is designed to work seamlessly with the Google ecosystem. Through the extensions feature, Gemini can connect with apps such as Gmail, Google Maps, YouTube, and Google Drive to help complete tasks more quickly. This capability allows Gemini, for example, to summarize email contents, search for travel routes, or find information from videos without having to open the apps separately. On Pixel 9 devices and later, Gemini is the default assistant, accessible by long-pressing the power button or using the voice command "Hey Google".

Personalization with Gems

One of the main personalization features is Gems. Gems are specialized versions of Gemini that can be customized to become "experts" on a particular topic without requiring coding skills. Users can create Gems for specific purposes such as writing editors, event planners, or sentiment analysis, by uploading instructions and reference files. Google also provides ready-made Gems, such as Learning coach for study guidance and Career guide for career preparation, which are highly useful in the education sector.

Gemini Live and Conversation

The Gemini Live feature delivers a more natural conversational experience. Users can speak with Gemini in real time, use it to brainstorm ideas, or simulate important conversations. Gemini Live is designed for smooth interaction, allowing users to share their screen or camera while speaking in order to receive contextual responses.

Models and Versions

Google releases Gemini in various model variants to meet different needs and budgets. Each model has optimized specifications, ranging from speed and cost to intelligence and multimodal capability.

Gemini 3.5 Flash

This model is a mainstay for tasks that require substantial intelligence at high speed. Released generally in May 2026, Gemini 3.5 Flash is known to be highly effective for programming tasks and long-term "agentic" tasks, offering performance comparable to large "flagship" models at a lower cost.

Gemini 3.1 Flash-Lite

This variant is optimized for speed, scale, and cost efficiency. The model is well suited to simple tasks and image generation with ultra-low latency.

Gemini Omni Flash

The newest model, still in public preview. Its main capability is conversational video creation and editing. Equipped with an invisible SynthID digital watermark for content verification, this model represents a new generation of generative AI capable of creating content from various types of input.

Image Models (Nano Banana)

Gemini also has specialized visual models such as Gemini 3.1 Flash Image ("Nano Banana 2") and Gemini 3 Pro Image ("Nano Banana Pro"). These models have been generally available and excel at image creation and editing, including a feature for generating images from video.

Availability and Subscriptions

Gemini can be accessed through various platforms, including mobile apps for Android (Android 10+) and iOS, as well as through the website at gemini.google.com. Google offers several subscription plans.

The Google AI Pro plan unlocks access to the most advanced models such as 2.5 Pro, the Deep Research feature for detailed reports, and eight-second video generation with Veo 3, as well as a 1-million-token context window capable of processing up to 1,500 pages of text.

The Google AI Ultra plan is the highest tier, offering exclusive access to the most powerful models, the latest features such as Agent Mode, and the opportunity to try Google's AI innovations earlier. This plan is available in certain countries and is not for Workspace customers.

In Southeast Asia, Gemini has recorded significant user growth, more than doubling in a year. Vietnam has become the regional leader in using Gemini for academic purposes and local-language use. The country has even become the market with the highest percentage of native-language (Vietnamese) use in the region, reaching 89%.

Impact and Education

Gemini, along with other generative AI, has begun to change the way people access information. This change has sparked discussion about the future of traditional knowledge sources such as Wikipedia. Wikimedia Indonesia, for example, held the #WikipediaxAI program in 2025 to discuss how AI uses Wikipedia content and how the community can adapt. Concerns have arisen that the shift from reading articles to receiving instant summaries from AI could reduce the number of visitors to source sites and the number of new volunteer contributors.

On the other hand, Gemini offers great potential in education. In Vietnam, more than 160,000 students use the Gemini Canvas feature every month for exam preparation, while educators generate tens of thousands of teaching assistance requests every day. This use shows AI's role as a powerful supporting tool, not a replacement for human critical thinking. A comparative study even placed Gemini as a tool comparable to ChatGPT for various translation tasks, demonstrating its capability in language processing.

References

1. Use Gemini, your AI assistant, on your Pixel phone. Google Help.

2. Education/News/December 2025/WikipediaxAI: Wikipedia, AI, and the future of knowledge. Meta-Wiki, Wikimedia.

3. Catatan rilis, Gemini API. Google AI for Developers.

4. Personalize Gemini with Gems: custom AI experts for any topic. Google Services.

5. Google Gemini, Apps on Google Play.

6. Cara Menggunakan Gemini AI Google Bahasa Indonesia, Mudah dan Praktis. Kompas.com.

7. Release notes, Gemini API. Google AI for Developers.

8. Gemini, "Asisten" Canggih untuk Sehari-hari. LinkedIn, 360 Solusi Teknologi.

9. Việt Nam dẫn đầu Đông Nam Á về dùng AI của Google hỗ trợ học thuật. VietnamPlus.

10. Digilib UIN SGD (Bab 1).

11. 100 things we announced at I/O 2026. Google Blog.

12. Try Gemini, your personal AI assistant. Android.com.

[Read more →]

Naver

Naver (Hangul: 네이버) is a South Korean web portal and search engine operated by Naver Corporation. Launched in 1999, Naver was the first web portal in South Korea to develop its own search engine and has transformed into one of the most influential technology companies in Asia, known not only for its dominant search engine but also for its extensive digital services ecosystem, including the instant messaging app Line and the web comic platform Webtoon. The name "Naver" itself comes from the word "navigate" with the suffix "-er", interpreted as "a navigator in the Web world".

History and Development

Founding and Early Years

Naver was founded in June 1999 by a group of former Samsung Data Systems employees led by Lee Hae-jin. Initially, the company was named Naver Comm with the goal of creating a Korean-language search engine capable of competing with foreign portals such as Yahoo, which at the time dominated the market. Naver's main advantage from the beginning was the development of a search engine optimized for the distinctive features of the Korean language, which has a syllable structure and grammar very different from Indo-European languages, making it more relevant for local users than foreign competitors.

In 2000, Naver launched the "Comprehensive Search" feature, which allowed search results from various categories such as blogs, websites, images, and news to be displayed on a single page, an innovation made five years before Google launched a similar feature.

Merger and Early Expansion

In July 2000, Naver merged with Hangame Communications Inc., a leading online game portal in South Korea founded by Kim Bum-su. This merger produced a new entity named NHN Corporation (Next Human Network) in 2001, which combined the strengths of a search engine and a game portal to create the largest internet company in South Korea.

In the same year, Naver also acquired Search Solutions, a company developed by Soongsil University professor Lee Jung-ho, which had advanced natural language search technology. This acquisition further strengthened Naver's search technology capabilities. In 2002, Naver launched the "Knowledge Search" (Knowledge iN) service, a user-based question-and-answer platform that allowed users to ask questions and receive answers from other users. This service became very popular and helped Naver build a very large Korean-language content base, becoming a forerunner of modern Q&A platforms such as Quora. Naver was officially listed on the KOSDAQ stock exchange in 2002 and later moved to KOSPI in November 2008.

Main Services and Features

Search Engine and Portal

Naver is the leading search engine in South Korea, controlling around 60% of the search market share and becoming the most frequently visited desktop site with a 78% share of all website visits in the country. More than 25 million people in Korea set Naver as their browser homepage. The search engine processes more than 2.3 billion search queries every day.

Digital Services Ecosystem

Naver has developed a very extensive digital services ecosystem covering various aspects of South Korean users' digital lives. In 2004, Naver launched book and local information services, followed by a blog service in 2005 and Webtoon (web comics) in 2006, which revolutionized the digital comics industry with a freemium business model that allowed readers to enjoy comics for free but pay for exclusive content. Other services include news, email, academic thesis search, a children's portal (Junior Naver), and the online donation platform Happybean, launched in 2005 as the first of its kind in the world. Happybean allows users to find information and donate to more than 20,000 civil society and social welfare organizations. On April 1, 2013, Naver launched News Stand, which gave each news organization the freedom to edit its own articles appearing on Naver.

Mobile App

The NAVER mobile app provides various features optimized for mobile devices with four main tabs: 'Home' for everyday information such as weather and stock prices, 'Clip' for short videos, 'Content' for news and reading, and 'Shopping' for a personalized shopping experience. The app is also equipped with the Green Dot AI search feature, which includes image-based search (Lens), music search, voice search, and location-based search. In November 2024, Naver introduced a new "Topic Feed" and "My" to improve content personalization and the management of user activity.

Latest Technological Innovations

Naver continues to innovate in artificial intelligence technology. The company developed HyperCLOVA X, a Korean-made large language model (LLM), and collaborates with various global technology partners. In 2024, Naver announced a partnership with Intel to build an AI ecosystem based on the "Gaudi" chip. The company also uses AI technology for various services such as AI Briefing for search, ADVoost for automatic ad optimization, and AI Shopping Guide, which provides a highly personalized shopping experience. Naver continues to invest in research and development, with overseas research centers including Naver Labs, and uses Nvidia GPU resources for AI development.

Global Expansion

Line and International Presence

One of Naver's greatest international successes is the launch of Line, an instant messaging app developed by Naver Japan in June 2011. Line was originally developed as an emergency communication tool after the major earthquake and tsunami in Japan, but quickly became one of the most popular messaging apps in Asia, including in Japan, Taiwan, Thailand, and Indonesia. Line's distinctive features are its use of diverse colorful stickers and an integrated service ecosystem covering payments, games, and delivery services. In 2019, Naver consolidated the Line business with Yahoo Japan through a joint venture with SoftBank, creating a larger digital giant in Asia.

Naver has expanded its global presence by establishing offices in various countries including Japan, the United States, France, China, Vietnam, Taiwan, Thailand, and Indonesia. In Indonesia, Naver Corporation is recorded as an investor in PT Elang Mahkota Teknologi Tbk (Emtek) through an investment worth Rp 9.29 trillion in 2021.

Acquisitions and Growth

In 2022, Naver acquired Poshmark, the largest secondhand goods marketplace in North America, with an acquisition value of 2.3441 trillion won. This acquisition marked Naver's strategic step in expanding the reach of its global e-commerce business.

Financial Performance

Naver Corporation has shown solid financial performance with sustainable growth. In 2017, the company's revenue reached 4.06 trillion won. In 2023, the company's operating profit reached 3,727 billion won in the second quarter with 10.9% growth, while sales reached 2.4 trillion won with 17.7% growth. In 2024, Naver recorded revenue of more than 10 trillion won, with operating profit increasing 32.9% to 1.9793 trillion won. In the second quarter of 2025, the company recorded a net profit of 4,974 billion won with revenue of 2.915 trillion won.

Naver's financial success reflects its dominance in the South Korean market and its ability to innovate and adapt to technological change, with a focus on integrating artificial intelligence to strengthen the platform and optimize advertising.

References

Wikipedia bahasa Indonesia. "Naver". id.wikipedia.org.

CNBC Indonesia. "Resmi Jadi Investor Baru Emtek, Siapa NAVER Corporation?". cnbcindonesia.com.

РБК Тренды. "Корейский Google: как Naver стал технологической империей". trends.rbc.ru.

今周刊. "在韓國Naver比Google還厲害!開發台日爆紅的LINE、專注整合AI技術,謝金河:南韓傳奇企業". businesstoday.com.tw.

財訊. "Google在韓國為何卡卡?謝金河直擊Naver總部親揭關鍵". wealth.com.tw.

NAVER HELP. "주제피드와 새로워진 마이를 만나 보세요!". help.naver.com.

NAVER Corporation. "Featured Services". navercorp.com.

百度百科. "NAVER". wapbaike.baidu.com.

NAVER Corporation. "Press Releases". navercorp.com.

[Read more →]

Claude Code

Claude Code is an artificial intelligence-based coding agent software developed by Anthropic. Unlike conventional programming assistants that only provide code suggestions, Claude Code operates as a command-line interface (CLI) tool that can independently read, analyze, and modify an entire codebase. The tool is designed to help developers at various stages of the software development cycle, from exploring a new project and debugging to writing tests and large-scale refactoring.

History and Development

Claude Code was introduced by Anthropic as part of the Claude model ecosystem to meet the need for a more autonomous programming tool. Its main development has been marked by the launch of support for various surfaces, such as a native extension for Visual Studio Code (VS Code) and a redesigned desktop interface. One important milestone was the introduction of the checkpoint feature, which allows developers to "rewind" to a previous state of the code, providing psychological safety when carrying out large tasks. In 2026, Anthropic launched the dynamic workflows feature, which significantly increased Claude Code's capacity to handle complex projects by running hundreds of sub-agents in parallel.

Core Capabilities and Features

Claude Code offers various features that change the way developers interact with their codebases.

Terminal- and IDE-Based Agent

Claude Code runs natively in the terminal and is integrated with various Integrated Development Environments (IDEs) such as VS Code, JetBrains, and a desktop application. Users can start a session with the simple command claude in their project directory. In this mode, Claude can read the repository, edit files, run commands, and ask for confirmation before performing destructive actions. The checkpoint feature ensures that every change made by Claude can be undone at any time.

Dynamic Workflows

The dynamic workflows feature is a flagship capability that allows Claude Code to break complex tasks into several sub-tasks and run them in parallel using tens to hundreds of sub-agents in a single session. This enables Claude to handle work that previously took weeks in only a few days. A published case example is the use of dynamic workflows to port the Bun runtime from Zig to Rust, producing around 750,000 lines of Rust code with a 99.8% test pass rate in eleven days.

Customization and Control

Developers can give Claude permanent instructions through a CLAUDE.md file located at the project root. This file contains coding standards, architectural decisions, and a code review checklist that Claude will read at the start of every session. In addition, teams can create skills to automate repeating workflows (for example, a /ship command) and use hooks to run shell commands before or after certain actions.

Model Context Protocol (MCP)

Claude Code supports the Model Context Protocol (MCP), an open standard that allows AI agents to connect to various external data sources. With MCP, Claude Code can read design documents in Google Drive, update tickets in Jira, or retrieve data from Slack, making it a control center integrated with the tools already used by development teams.

Common Use Cases

Based on Anthropic's official documentation, there are ten main scenarios in which developers most often use Claude Code:

Fixing Failing Tests. When a test fails and the cause is unclear, Claude Code can trace the root problem and propose a fix without the developer first having to identify the source file.

Understanding Unfamiliar Code. Claude can guide developers through how a particular module or function works, explaining control flow and decision points in natural language.

Finding the Location of a Function. Developers can ask where a particular validation occurs, and Claude will return the exact file path and line number.

Classifying Errors. By providing a stack trace or error log, Claude can map it back to the responsible code and explain the cause.

Refactoring with a Plan. Before making large changes across many files, Claude can create a detailed plan of which files will be changed and how they will change, which the developer must approve first.

Writing Tests for Existing Code. Claude can generate new test files that match the project's existing coding style and run them to ensure everything works.

Reviewing Pull Requests. Integrated with the GitHub CLI, Claude can retrieve the diff, review comments, and CI status to produce a PR review or summary.

Onboarding to a New Repository. The /init command allows Claude to scan the project and generate a CLAUDE.md file that summarizes the architecture, build commands, and conventions used.

Handling Issues End-to-End. With MCP connectors, Claude can read a ticket, implement a fix, and validate it in a single conversation without having to switch tools.

Turning Repeating Tasks into Skills. Developers can create custom commands (such as /ship) that run a series of automated steps (for example, running tests, a linter, and creating a commit message).

Models and Pricing

Claude Code uses Anthropic's latest AI models, including Claude 3.7 Sonnet and Claude 4 Opus. It is available in several subscription plans, ranging from Claude Pro for individual use to Enterprise plans for large organizations. Team and Enterprise plans provide user management features, SSO, and usage analytics. Companies also have the option to run Claude Code through their own cloud infrastructure such as Amazon Bedrock, Google Cloud Vertex AI, or Microsoft Foundry, ensuring that data remains under their control and meets compliance standards such as SOC 2 Type II.

References

· claude.com, "Claude Code for Enterprise"

· claude.com, "Claude Code: Common developer use cases"

· claude.com, "Introducing dynamic workflows in Claude Code"

· claude.com, "Overview - Claude Code Docs"

· claude.com, "Enabling Claude Code to work more autonomously"

· IT Brief India, "Anthropic launches dynamic workflows in Claude Code"

· Media Indonesia, "Panduan Claude Code Anthropic 2026: Fitur, Harga, & Cara Pakai"

[Read more →]

Nano Banana

Nano Banana is the code name for an advanced artificial intelligence (AI)–based image generation and editing model developed by Google. Introduced in late August 2025 as part of the Gemini model family, specifically Gemini 2.5 Flash Image, the technology quickly became a viral phenomenon on social media because of its ability to generate and edit images with high quality, precision, and exceptional character consistency using only natural language prompts. The name "Nano Banana" itself comes from an efficient Gemini model variant ("Nano") and an unusual internal nickname during development ("Banana").

History and Development

Speculation about the existence of "Nano Banana" began circulating among internet users after Google engineers and eventually CEO Sundar Pichai posted banana emojis on social media. The mystery was answered on August 26, 2025, when Google officially announced the launch of its latest image model, Gemini 2.5 Flash Image, whose code name is Nano Banana. This launch marked a significant improvement over the previous generation image model, Gemini 2.0 Flash, with a focus on higher image quality and stronger creative control for users.

Since its launch, Nano Banana has experienced a spectacular surge in popularity. In its first month alone, more than 500 million images were generated by users around the world using its visual editing feature. Thanks to this model, the Gemini app succeeded in attracting 23 million new users and rose to the top of the App Store and Google Play in various countries, including Indonesia.

Core Capabilities and Features

Image Generation and Editing with High Consistency

Nano Banana's main advantage lies in its ability to produce unmatched image consistency. Unlike many other AI generators that often struggle to maintain visual continuity, Nano Banana excels at preserving intricate details such as facial expressions, hand movements, lighting, textures, and spatial relationships across generated elements. This capability is especially important for projects that require a uniform appearance of characters, style, or environmental details.

Users can upload an existing image and give natural language text instructions to edit it, such as changing the background, replacing clothing, or adding objects, while the model ensures that the main subject remains recognizable and consistent. For example, someone can upload a photo of themselves and see how they would look with a different haircut or particular clothing, all while still looking like themselves.

Natural Language Understanding and Conversational Editing

Nano Banana is designed to understand complex instructions in everyday language. Users do not need technical expertise such as using Photoshop; they simply describe what they want, for example "replace the cloudy sky with a sunset and add a flying drone," and the model will execute it. The conversational editing feature allows users to carry out a series of edits in sequence while keeping the image style and lighting consistent.

Multi-Image Merging

Another innovative capability is multi-image merging, which allows users to combine several images into one cohesive visual. Google gave the example that someone can upload a photo of themselves and a photo of their dog to create "the perfect portrait of the two of you on a basketball court." This fusion technology creates a perfect composition with a simple prompt.

Models in the Nano Banana Family

The platform offers several model variants tailored to different needs, all built on Gemini's multimodal capabilities and responsive even to the most detailed prompts.

Nano Banana Pro is the flagship model designed for image generation and editing with professional studio-level precision and control. This model is built on Gemini 3 and allows users to imagine and create almost anything, from detailed images to realistic or imaginative ones. Nano Banana 2 offers professional-level image generation and editing at lightning speed, based on the Gemini 3.1 Flash Image model. Meanwhile, Nano Banana 2 Lite is the fastest and most efficient model, providing the highest speed and lowest cost for generation and editing.

Impact and Viral Phenomenon

3D Action Figurine Trend

Nano Banana became highly viral thanks to the trend of turning ordinary photos into highly realistic 3D action figurines. Users used prompts to turn self-portraits, celebrities, or cartoon characters into collectible statues that look like toys in packaging, often with bases and custom packaging that make them look like products on a store shelf. This trend flooded social media platforms such as X, Reddit, and Instagram.

Democratization of Aesthetics and Democratization of Creativity

The Nano Banana phenomenon marks a kind of democratization of aesthetics, in which anyone can look like a celebrity or create professional-class visuals without needing design skills or expensive equipment. In Indonesia, this culture was quickly adopted, where people from various backgrounds can turn simple photos into stunning works of art.

Cross-Platform Integration

Google has aggressively integrated Nano Banana into its application ecosystem. After its success in the Gemini App and Google AI Studio, the model is in the process of being integrated into Google Messages, allowing users to edit and generate images directly from text conversations. It has also been demonstrated in Google Lens and is being developed for Google Photos.

Access and Availability

Nano Banana can be accessed globally through the Gemini app or Google AI Studio. The model is available to free and paid users, although free users may have daily production limits and slower processing speeds, while paid accounts remove most of those limitations.

All images created or edited with Nano Banana include a SynthID digital watermark that is both visible and invisible to indicate that the image was generated by AI.

Impact and Criticism

Although it offers unlimited ease and creativity, the emergence of Nano Banana has also raised questions about authenticity and psychological impact. The ability to drastically change appearance so easily can lead to cultural dissonance, in which people trust their digital selves more than their real selves. There are concerns that a generation of users may become more occupied with arranging AI prompts than living real life.

The question of "who we are without filters, without edits, without a nano banana overlaying our faces" has become an important reflection amid the onslaught of technology that increasingly blurs the line between reality and virtuality.

References

Google DeepMind. "Nano Banana". Google DeepMind. Diakses pada 15 Juli 2026.

Media.io. "Nano Banana AI: Apakah Ini Akhir dari Pengeditan Gambar Tradisional? Inilah Cara Menguasainya". Media.io. 6 September 2025.

Hindustan Times. "'Nano Banana' is taking over internet: All about viral AI action figure craze". Hindustan Times. 11 September 2025.

Trusted Reviews. "What is Nano Banana? Google’s mysterious technology is finally unveiled". Trusted Reviews. 28 Agustus 2025.

Duta.co. "Pribadi Nano Banana". Duta.co Berita Harian Terkini. 18 September 2025.

Radar Banyuwangi. "Viral! Tren Sulap Foto Lama Jadi Kekinian Pakai Gemini AI, Gampang Banget!". Radar Banyuwangi. 25 September 2025.

Alibaba.com Reads. "Bersedia untuk Kejutan 'Pisang Nano' dalam Mesej Google". Alibaba.com Reads. 19 Oktober 2025.

TEGAROOM ONLINE. "Mengubah Imajinasi Menjadi Visual: Google Perkenalkan Nano Banana 2025". 30 November 2025.

[Read more →]

Gemini 3

Gemini 3 is the latest generation of multimodal artificial intelligence (AI) models developed by Google DeepMind. Announced on November 18, 2025, this model is the successor to Gemini 2.5 and is described by Google as their "most intelligent" model as well as "the world's best for multimodal understanding". This launch marked a significant step in Google's effort to lead the generative AI race, with the model integrated directly into various of their flagship products, such as Google Search and the Gemini app, on the same day of launch.

Core Capabilities

Reasoning and Multimodal Understanding

Gemini 3 is built with state-of-the-art reasoning capabilities designed to understand the depth, nuance, and context behind user requests. Unlike its predecessors, which often required explicit format instructions, Gemini 3 introduces a "generative interface". This feature allows the model to determine the best output format itself, such as creating an immersive visual layout, a diagram, or even a simple animation, if it is considered more effective than ordinary text. The model also shows high performance on various benchmarks, including a score of 1501 Elo on the LMArena leaderboard, and a score of 37.5% on the Humanity's Last Exam test that measures PhD-level reasoning.

Agentic Capabilities and "Vibe Coding"

One of the main improvements in Gemini 3 is its capability as an agentic AI. The model can not only respond to questions, but also plan and execute multi-step tasks independently. This capability is strengthened by the introduction of Google Antigravity, a new development platform that allows developers to delegate complex tasks, such as writing, testing, and verifying code, to AI agents that can work across the editor, terminal, and browser. This approach of enabling the creation of an entire interface or application from a single natural language prompt is known as "vibe coding".

Model Variants

Gemini 3 Pro

Gemini 3 Pro is the flagship variant designed to handle complex tasks, deep reasoning, and creative content generation. This model is available to developers through Google AI Studio and Vertex AI, as well as to Google Cloud customers. Pro shows significant improvements in visual understanding, code generation, and performance on long-running tasks. The model also serves as the foundation for various new features in Google Search and the Gemini app.

Gemini 3 Flash

Gemini 3 Flash is a faster and more efficient variant, designed for high-frequency tasks that require greater speed and lower cost. Although lighter, Flash still offers academic-level reasoning capabilities and excels at agentic coding tasks, with a score of 78% on the SWE-bench Verified benchmark. This model is the standard model in the Gemini app for everyday use.

Gemini 3 Deep Think

The Gemini 3 Deep Think feature is a special reasoning mode that pushes the model's capabilities further by exploring several hypotheses in parallel to solve highly complex problems, such as scientific research, advanced mathematics, or intricate planning. This mode was first made available to Google AI Ultra subscribers.

Impact and Reach

With the launch of Gemini 3, Google emphasized a "Google scale" approach by bringing the model simultaneously to various platforms. This includes integration in Google Search, the Gemini app, developer platforms such as AI Studio and Vertex AI, and the new Antigravity platform. The company also offered one year of free access to Gemini Pro for students in the United States to encourage adoption in academic circles. This step reflects Google's strategy of not only developing advanced technology, but also ensuring broad accessibility and integration in everyday life and the professional world.

References

· MIT Technology Review. Google’s new Gemini 3 “vibe-codes” responses and comes with its own agent.

· Google Cloud Blog. Gemini 3 is available for enterprise.

· Bloomberg. Google Launches New Gemini AI Model With Interactive Answers.

· Google Blog. Gemini 3 Membuka Era Kecerdasan Baru.

· CNN Indonesia. Google Rilis Model AI Gemini 3, Bisa Apa Saja?

[Read more →]

ChatGPT

ChatGPT is an artificial intelligence (AI) model developed by OpenAI, designed to generate text similar to human writing based on input provided by the user. Launched in November 2022, ChatGPT quickly became a global phenomenon and changed the way society interacts with technology, from professional work to everyday personal use. Its presence has not only revolutionized productivity but also sparked in-depth discussion about the future of knowledge, education, and social interaction in the digital era.

History and Development

The technology underlying ChatGPT is OpenAI's GPT (Generative Pre-trained Transformer) model series, which has undergone a number of significant improvements. Since its debut with GPT-3.5, the model has transformed through various versions up to GPT-5.5, which has become the most frequently used model. Each new generation has brought improvements in reasoning ability, following of complex instructions, and overall conversation quality.

This development marks a shift from merely a language model to a more intelligent assistant. In mid-2026, models such as GPT-4.5 and GPT-5.2 were officially retired in ChatGPT and shifted to newer versions to ensure users receive the best experience. In addition, OpenAI periodically releases new models such as GPT-5.5 Instant Mini, which functions as a backup model when users reach usage rate limits, ensuring continued access to the service.

Main Features and Capabilities

ChatGPT has evolved from merely a chat box into an advanced platform with a variety of features designed to increase productivity and personalization.

Context and Personalization

One of ChatGPT's main capabilities is memory and personalization. This feature allows ChatGPT to remember information from previous conversations, providing more relevant and contextual responses over time. Users can view a memory summary, correct stored information, or turn off the memory function entirely to maintain privacy. This personalization makes ChatGPT an assistant that is increasingly adaptive to each user's specific needs.

Scheduled Tasks

The "Scheduled Tasks" feature allows users to ask ChatGPT to perform repeating work, send reminders, or monitor something at a predetermined time. This changes ChatGPT from a reactive tool into a proactive tool that can be relied on for various daily routine needs.

Codex and Work Automation

The development of Codex marked a major leap in ChatGPT's operational capabilities. Features such as Codex Remote allow users to control a remote computer from a phone, while Record & Replay allows users to demonstrate a workflow and turn it into a reusable skill. For business users, this opens large opportunities to automate repeating tasks in various fields, such as software development, data analysis, and project management. The Developer mode for Browser use feature also gives developers more control for debugging and performance analysis.

Productivity and Accessibility Improvements

In addition to the features above, ChatGPT has consistently introduced improvements to the user experience. Improvements to the dictation model with a word error rate lower by up to 10% make it easier for users to interact without typing. For users who enter long text, ChatGPT now automatically converts pastes of more than 10,000 characters into attachments, keeping the chat box tidy. Pronunciation assistance features for more than 60 languages and the integration of current information updates such as World Cup schedules also add to ChatGPT's usefulness.

Impact on Knowledge and Education

The presence of ChatGPT has had a major impact on the way humans access and process information. As a language model trained on large amounts of data from the internet, including Wikipedia articles, ChatGPT is able to provide instant and concise answers to user questions. This has changed information-seeking behavior, in which users can obtain knowledge without having to visit the source site directly, such as Wikipedia, which is one of its main training data sources. A study even showed that Wikipedia is the source most frequently cited by ChatGPT, with a share reaching 16.3% of all citations.

Role in Education

In the world of education, empirical studies have explored ChatGPT's potential as a cognitive scaffold. Research in 2026 at a madrasah in Indonesia showed that the use of ChatGPT in argumentative writing instruction had a large effect on student achievement, particularly in the aspects of organization and written content. ChatGPT was shown to help students construct a framework of thinking and develop arguments, although its effectiveness was not even across all linguistic aspects such as vocabulary and grammar. Another study also looked at the use of ChatGPT in ecology-based Indonesian language learning, opening opportunities for more innovative teaching methods.

Knowledge Dynamics in the AI Era

Experts see that generative AI, including ChatGPT, will change the knowledge landscape. There are concerns that if AI becomes increasingly advanced, the role of encyclopedias such as Wikipedia as a center of knowledge could be displaced. However, there is also a view that AI can instead become a tool for democratizing knowledge, by translating and disseminating information into various regional languages that were previously difficult to access. This discussion highlights the importance of the role of human curators in ensuring the quality and validity of information, even if that content is generated by AI.

Regulation and Governance

Along with its growing popularity, ChatGPT and other digital platforms have faced oversight from regulators in various countries. In Indonesia, in November 2025, the Ministry of Communication and Digital Affairs (Komdigi) issued a warning to dozens of major digital platforms, including OpenAI (ChatGPT) and Wikipedia, because they had not yet registered as Private-Scope Electronic System Operators (PSE). This step was taken to ensure compliance with Minister of Communication and Informatics Regulation Number 5 of 2020, with the threat of sanctions in the form of access blocking if not fulfilled immediately.

Social Phenomenon: Confiding in AI

The phenomenon of ChatGPT use is not limited to a productivity tool, but has also entered the social-emotional realm. Digital sociology research in late 2025 revealed that many users in Indonesia use ChatGPT as a "confidant." The main factors behind this phenomenon are the perception that AI is non-judgmental, always available 24 hours a day, provides emotional validation, and offers a sense of safety in terms of privacy. However, this research also warned of potential negative impacts, such as emotional dependence, widening social distance, and the risk of data leaks. This highlights the need for balanced digital literacy so that AI does not replace social relations among humans.

References

· Wikimedia. (2025, Desember 15). Education/News/December 2025/WikipediaxAI: Wikipedia, AI, and the future of knowledge. Meta-Wiki. 

· OpenAI Help Center. (2026, Juni 26). ChatGPT — 版本说明. 

· OpenAI Help Center. (2026, Juli 6). ChatGPT — Release Notes. 

· OpenAI Help Center. (2026, Juni 26). ChatGPT Business - 发行说明. 

· OpenAI Help Center. (2026, Juni 26). ChatGPT Business - 發布說明. 

· Hops.ID. (2025, November 19). Cek HP Anda Sekarang! 25 Aplikasi Favorit Ini Terancam Hilang, Ada ChatGPT hingga Wikipedia. 

· Bennylin. (2023, Mei). Pengguna:Bennylin/20TahunWikipedia. Wikipedia bahasa Indonesia. 

· Syafawani, R., dkk. (2025, Desember 23). Fenomena Curhat ke AI: Tinjauan Sosiologi Digital Terhadap Pergeseran Interaksi Sosial di Indonesia. Jurnal Sosial dan Sains. 

· Jumaddil, J., dkk. (2026, April 30). ChatGPT as a Macro-Level Cognitive Scaffold for First-Language Argumentative Writing. Edcomtech. 

· Steavenson, M. (2025, Juni 14). AI Search Visibility: Why Reddit, Wikipedia, and YouTube Win And What It Means For Your Brand. LinkedIn. 

· Abdian Amien, A., & Kusumawati, H. (2024, Desember). Optimalisasi CHATGPT Dalam Pembelajaran Bahasa Indonesia Berbasis Ekologi Di SMAN 4 Pamekasan. GHANCARAN. 

[Read more →]

Google

Google LLC is a multinational technology company from the United States known as a provider of internet services and products, especially the world's dominant search engine. It was founded on September 4, 1998 by Larry Page and Sergey Brin when both were still doctoral students at Stanford University. The name "Google" itself was inspired by the mathematical term "googol" (the number 1 followed by 100 zeros), which reflects the company's mission to organize the world's information in enormous quantities and make it universally accessible.

The company officially became part of the conglomerate Alphabet Inc. in 2015 as a result of a corporate restructuring. Over time, Google has evolved from merely a search engine into a technology giant with a very broad product portfolio, covering software, hardware, and various artificial intelligence (AI)–based services used by billions of people around the world.

History and Early Development

Beginnings and Founding

In 1995, Larry Page and Sergey Brin met at Stanford University when Page was considering continuing his studies at the university and Brin was assigned to show him around campus. Although they often disagreed at first, they later developed a shared interest in information retrieval systems on the internet. Their research project began with a search engine named BackRub, named after its ability to analyze "backlinks" in order to determine the importance of a web page.

The two developed an innovative algorithm that later became the foundation of Google's technology, known as PageRank. This algorithm was able to rank search results based on the number and quality of links pointing to a page, an approach that proved far more relevant than other search methods of the time. In 1997, they officially renamed their search engine Google. The company Google LLC was then officially founded the following year after Andy Bechtolsheim, one of the co-founders of Sun Microsystems, became interested in their demonstration and wrote a check worth 100,000 U.S. dollars as initial capital.

Growth into a Technology Giant

Google Search quickly became the most widely used search engine in the world, surpassing competitors such as Yahoo! and AltaVista. This success was driven by more accurate search results and a clean, simple interface. By the mid-2000s, Google had become one of the most influential technology companies. In 2004, Google launched the free email service Gmail, which offered very large storage capacity and significantly changed the standard of email services.

Google's expansion became increasingly aggressive with a series of strategic acquisitions. In 2005, they acquired the mobile operating system developer company Android Inc., which would later become the foundation of the world's most popular operating system. Not long after, in 2006, Google acquired YouTube, a video-sharing platform that had become a global phenomenon and one of the most visited sites on the internet. These acquisitions strengthened Google's dominance not only in search, but also in digital advertising, video content, and the mobile device ecosystem.

Innovation and Main Products

Search Engine (Google Search)

Google's flagship product to this day is its search engine. With a global market share reaching more than 90%, Google Search is the main gateway for most internet users to access information. For more than 25 years, the Google search box has been the company's primary symbol.

At the annual Google I/O 2026 developer event, the company announced the largest update to their search box since it was first launched. This new search box is powered by artificial intelligence and designed to process longer, more natural, and more complex questions. Users can now search not only with text, but also with images, files, videos, and even currently open browser tabs. Search results are also no longer limited to a list of links, but can take the form of contextual answers summarized by AI, interactive graphs, or even dynamically created user interfaces according to the question asked.

The Android Operating System

Android is a Linux kernel–based mobile operating system acquired by Google in 2005. Since its launch, Android has become the backbone of the mobile device ecosystem outside Apple's environment. Its presence has allowed Google to bring its services, such as Search, Gmail, and Google Maps, to billions of smartphone users around the world. Android is also known as one of Google's most widely used products in Indonesia.

Other Popular Products

In addition to Search and Android, Google has various products with more than one billion users. YouTube, acquired in 2006, has become the world's largest video site. Gmail revolutionized email with large storage and advanced features. Google Maps is the most popular digital map application, providing navigation and location information in real time. The Google Chrome browser also dominates the market with its speed and extensive extensions. The photo and video storage service Google Photos has also been used by billions of people to back up and manage their digital memories.

Artificial Intelligence: The Gemini Model and the Future

Gemini: Flagship AI Model

In recent years, Google has shifted its main focus to the development of artificial intelligence, with their flagship model named Gemini. Google's long journey in AI began in the Google Brain era, which in 2011 started the Distbelief project and produced the famous "Cat Paper", which showed that a deep neural network could learn to recognize objects such as cats from YouTube videos without labels. Its peak innovation came in 2017 when the Google Brain team published a paper on the Transformer architecture, which became the foundation for the modern AI revolution and gave rise to models such as ChatGPT and Gemini itself.

At Google I/O 2026, Google launched Gemini 3.5 Flash, the latest model in the Gemini series that combines advanced intelligence with high speed. This model is generally available and designed to handle "agentic" tasks or tasks that require a series of complex actions, such as helping developers create applications or prepare financial documents. Google also introduced Gemini Omni, a revolutionary model that can create various types of content from various types of input (text, images, video, audio), with an initial focus on video creation.

AI Agents and Cross-Platform Integration

Google introduced various AI agents designed to help users in everyday life. Gemini Spark is a personal AI agent that can work in the background around the clock to help manage users' digital lives, including reading email, summarizing documents, and taking action on the user's behalf after receiving approval. The Daily Brief feature arrives as an agent that automatically gathers information from calendars, email, and tasks to provide a personal summary in the morning. These agents demonstrate Google's vision of making AI a new operational layer throughout all of their products, from Search and Gmail to YouTube.

Google's Presence in Indonesia

Establishment of a Branch Office and Tax Compliance

Google officially opened its first branch office in Indonesia on March 30, 2012, making Indonesia the fourth country in Southeast Asia to have a Google office after Singapore, Malaysia, and Thailand. This establishment was driven by large market potential, given that in the same year, nearly 80.91 percent of internet users in Indonesia used Google as their main search engine.

Although it already had an office, there was a tug-of-war between Google and Indonesia's Directorate General of Taxes (DJP) regarding tax obligations. This culminated in the issuance of Minister of Finance Regulation Number 35/PMK.03/2019 on the Determination of a Permanent Establishment (BUT), which pushed Google Asia Pacific Pte. Ltd., operating in Singapore, to establish a BUT in Indonesia. Finally, in 2019, Google Asia Pacific Pte. Ltd. was willing to carry out its tax obligations in Indonesia, including paying Value Added Tax (PPN) and Income Tax (PPh).

Contribution to the Indonesian-Language Wikipedia

Google also collaborated with the Wikimedia Foundation and Wikimedia Indonesia in a pilot project that began in 2019. This project aimed to expand access to Indonesian-language information on the Google search engine. In this project, when a user searches for a topic in Indonesian and there is no quality Wikipedia article in that language, Google will display a machine-translated version of the English Wikipedia article. This result will be clearly labeled as a machine translation, and if there is a quality Indonesian-language article, Google's algorithm will prioritize that original article. This collaboration is part of a broader project such as Project Gayatri, which aims to support an increase in the quantity and quality of content on the Indonesian-language Wikipedia. This initiative highlights a joint effort to address the content gap, in which the English Wikipedia has tens of millions of articles compared with fewer than one million articles on the Indonesian-language Wikipedia.

References

Google I/O 2026: all our announcements. (2026). Google Blog.

Global Reach/Announcements/Bahasa Indonesia Search Pilot FAQ. (2019). Wikimedia Meta-Wiki.

Google AI 编年史:从搜索巨头到创新者困境的25年. (2025). 36氪.

Indonesia Gen Z: Data, Demographics, and Digital Behavior. (2026). Jakpat.

Seperempat Abad Google dan Aspek Perpajakannya sebagai BUT di Indonesia. (2023). Direktorat Jenderal Pajak.

Wikipedia bahasa Indonesia. (2025). Wikipedia.

Seperempat Abad Google dan Aspek Perpajakannya sebagai BUT di Indonesia. (2023). Direktorat Jenderal Pajak.

Google I/O 2026: Tất cả công bố đáng chú ý. (2026). VTV.

Catch up on 12 major I/O 2026 moments. (2026). Google Blog.

Proyek Gayatri Overview. (n.d.). Wikimedia Outreach Dashboard.

Google 滿 20 歲了:它如何改變網路,撼動世界?. (2018). INSIDE.

Tepat Hari Ini, 20 Tahun Silam Google Dibentuk hingga Akhirnya Jadi 'Raja Dunia'. (2018). Serambinews.

[Read more →]