Skip to content
MakataLet's talk

Building the foundationsof Philippine intelligence.

Makata is an independent Filipino AI research initiative exploring language, local knowledge, and the foundations of intelligence we can help build ourselves.

Filipino-first. Culturally-aware

Why Makata exists

A place for our languages.
A role in our own future.

Language carries more than information. It carries how we think, the things we know, and the context that makes us understood. Philippine languages deserve a deeper place in AI research.

We want to help the Philippines build the capability to create, study, and operate AI—not only use it. That begins with patient, independent work, informed by the wider research community.

Representation

More room for Filipino, Taglish, and the language we use in everyday life.

Research capacity

Local knowledge and engineering skills to investigate the models behind the applications.

Greater independence

The long-term ability to make informed choices about our data, models, and infrastructure.

Research directions

Small beginnings.
Foundational questions.

Our initial focus is Filipino and Taglish. These are the areas we intend to investigate as our research capabilities develop.

Philippine-language models

Investigating how language models understand, generate, and reason in Filipino and Taglish, including the ways we move between languages.

Language datasets

Exploring the collection, curation, and documentation of Philippine-language data, with careful attention to provenance, privacy, and licensing.

Training & adaptation

Studying open-weight models and controlled adaptation experiments, while building the knowledge needed for original model research.

Independent evaluation

Planning reproducible evaluations that make capabilities, limitations, and performance in Philippine contexts visible.

Efficient, local AI

Investigating smaller models and efficient approaches that could run on accessible hardware and locally controlled infrastructure.

Other Philippine languages are a longer-term direction, guided by suitable datasets and language expertise.

Our approach

Work that can
be examined.

Credibility comes from how the research is done—and what we are willing to make clear.

makataFilipino noun · poet

A name rooted in language, expression, and the knowledge we carry.

Reproducible by design

Document methods, configurations, and limitations so that others can understand and eventually repeat our experiments.

Language in context

Treat linguistic nuance, code-switching, and cultural knowledge as central research questions.

Evidence before claims

Evaluate against clear baselines. Report regressions alongside improvements, and distinguish plans from demonstrated results.

Responsible foundations

Respect data rights and privacy. Be explicit about model origins and the difference between adaptation and training from scratch.

Where we are

At the beginning,
with a clear direction.

Foundational stage

We are establishing Makata's research environment, documentation, and public identity.

Next comes defining manageable baseline experiments, evaluating candidate open-weight models, and investigating appropriately sourced data.

No Makata model has been trained or publicly released. Baseline evaluations and model adaptation are planned work.

An open invitation

Good research begins
with a conversation.

Researchers, language specialists, engineers, universities, and institutions: if these questions matter to you, we'd like to connect.

Explore a collaboration jansencadorna5@gmail.com