hudsonbay

hudsonbay

Hi, please. I have a question regarding migrations in Ecto. But maybe this more of a PostgreSQL question rather than an Elixir one.

I have a many-to-many relation between students and teachers. One student can have many teachers and the same teacher can have many students.

So, I’m basically defining my migration like this:

 def change do
    create table(:students_teachers, primary_key: false) do
      add :id, :binary_id, primary_key: true
      add :criteria, :integer, null: false

      add(:student_id, references(:students, on_delete: :delete_all, type: :binary_id),
        null: false
      )

      add(:teacher_id, references(:teachers, on_delete: :delete_all, type: :binary_id),
        null: false
      )
    end

    create index(:students_teachers, [:student_id])
    create index(:students_teachers, [:teacher_id])
  end

Ok, my doubt is about the last line of the code, the part where I define the indexes:

create index(:students_teachers, [:student_id])
create index(:students_teachers, [:teacher_id])

But if I do it this in a second way create index(:students_teachers, [:student_id, :teacher_id]) something is gonna change.

If I do a diff to the two \d students_teachers generated tables in PostgreSQL with the different types of migrations I will see that there are some changes.

This the output of my diff (which BTW only sees the difference in the index definition):

>     "students_teachers_student_id_index" btree (student_id)
>     "students_teachers_teacher_id_index" btree (teacher_id)
---
<     "students_teachers_student_id_teacher_id" btree (student_id, teacher_id)

So, my question is:
How do this two approaches change the way my data is related? What is the difference here? How do the teacher and the student are related according to the different ways I migrated the database?

Showing Posts 1 to 7

cenotaph

cenotaph

Individual indexes on student_id and teacher_id fields

SELECT ..... WHERE student_id=1 => FAST SELECT
SELECT ..... WHERE teacher_id=1 => FAST SELECT

You would be telling Ecto to create a composite index

SELECT ..... WHERE student_id=1 => SLOW SELECT
SELECT ..... WHERE teacher_id=1 => SLOW SELECT
SELECT ..... WHERE student_id= AND teacher_id=1 => FAST SELECT

yurko

yurko

The first two queries are not that bad either, they’d still use indexes, especially the first one. Here’s some info about that:

hudsonbay

hudsonbay OP

Yes, but than can also be solved if I pass the :id of the table (students_teachers_id). Then, by using Repo.preload I can have the student’s data and the teacher’s data, right?. With that said, I still have doubts on what on what is the best thing to do regarding on what type of index to use (:

hudsonbay

hudsonbay OP

So what you are saying is that if I use this create index(:students_teachers, [:student_id, :teacher_id]) I can have this:

SELECT ..... WHERE student_id=1 => FAST SELECT
SELECT ..... WHERE teacher_id=1 => FAST SELECT
SELECT ..... WHERE student_id= AND teacher_id=1 => FAST SELECT

?

cenotaph

cenotaph

Let’s focus on your SQL question first

Rule of thumb:

  • If you need to search and access the records individually by teacher_id OR student_id you should create two indexes individually. I assume this is 99% the use case for most tables.
  • If you ALWAYS fetch the records from that table by teacher_id AND student_id, you should create a composite index.

Nope, that was not my statement. My statement was SLOW SELECT on top 2.

For clarification, when I say SLOW it is relative to correct index scan. SQL storages are fast enough to make up for the missing time, it is not easy for us humans to grasp the gap of milliseconds, microseconds, nanoseconds.

But that nano, milliseconds add up to minutes, hours, days over millions, billions of queries, transactions.

Repo.preload is all related to your definition of relationships. As long as there is a foreign_key (relation) path for Ecto to figure out the relationship, it will load them for you, regardless of how slow or inefficient the relationship is.

hudsonbay

hudsonbay OP

Ok, I get it. thanks

yurko

yurko

If you want to simplify the answer and make it slow / fast, then with your index (student_id, teacher_id) it would be

SELECT ..... WHERE student_id=1 => FAST SELECT
SELECT ..... WHERE teacher_id=1 => SLOW SELECT
SELECT ..... WHERE student_id= AND teacher_id=1 => FAST SELECT

See the link in my above comment for more on the topic of how slow the second query would be (not that slow).

From the Postgres docs (PostgreSQL: Documentation: 9.6: Multicolumn Indexes):

A multicolumn B-tree index can be used with query conditions that involve any subset of the index’s columns, but the index is most efficient when there are constraints on the leading (leftmost) columns.

— All posts loaded —

Where Next? Top

Trending in Questions Top

RSP87
I’m working on a project that simulates the bumbl example in the programming phoenix book. It acts almost like an email client. We have a...
New
kszambelanczyk
Hello! Could someone please give me a help/sample code, how to delete a file from s3 using waffle/waffle_ecto from Phoenix app. I creat...
New
RemyXRenard
I’m seeing that a list inside a Kino.DataTable will be interpreted as a charlist, even if the Kino.configure() is set to charlists: :as_l...
New
velrest
So my question is quite simple and i have found no conclusive answer on forum, google or AI. Should we use :erlang.float for Integer to ...
New
samoloth
Hi, I’ve just set up an application with ash_authentication. There is only magic link strategy for now, so there is no confirmation add o...
New
FlyingNoodle
If a change or preparation module uses Ash.Changeset.get_argument/2 or Ash.Query.get_argument/2 (or any of the other get_argument functio...
New
ryanwinchester
apply_graft/2 doesn’t rewrite an add_many sub-workflow’s deps on an add step. Grafted jobs cancel with “upstream job was deleted” Version...
New

Other Trending Topics Top

mudasobwa
I am happy to introduce the very α version of the new programming language compiled to BEAM. Welcome Cure. It has literally three kille...
New
marciok
Hi there! We created Gust: A task orchestrator inspired by Airflow. For those who have never heard about Aiflow, it’s a Python-based wor...
New
jimsynz
Beam Bots (or just BB for short) is a framework for building fault-tolerant robotics applications in Elixir using familiar OTP patterns. ...
New
Dmk
Xamal is a deployment tool for Elixir apps that deploys native releases to bare metal servers over SSH. It’s a port of GitHub - basecamp/...
New
Damirados
Hello everyone. After busy few months I am happy to announce v0.1.0 of Emerge &amp; Solve. They are GUI (Emerge) and State management (S...
New
netoum
Corex is an accessible, unstyled UI component library for Phoenix that integrates Zag.js state machines using Vanilla JavaScript and Live...
New

We're in Beta

About us Mission Statement

Options

Thread Display Mode




Thread Preview

Skip Thread Previews