travisf
I have a web app that handles a fair amount of traffic and sends a good deal of transactional emails, likely around 1000 a day. About 99% of the time there are no issues but about three months ago we started having an issue where emails are failing with this error:
{%Bamboo.SMTPAdapter.SMTPError{
message: "There was a problem sending the email through SMTP.\n\nThe error is :no_more_hosts\n\nMore detail below:\n\n{:permanent_failure, 'port', :auth_failed}\n",
raw: {:no_more_hosts, {:permanent_failure, 'port', :auth_failed}}
},
We get this error between 0 - 20 times a day but, in almost all cases, manually sending the email again solves the problem. Again, this application has been in production, and processing about the same number of orders, since 2019 and we only started seeing this issue in April or May of this year. Any ideas?
Trending in Questions
I’m working on a project that simulates the bumbl example in the programming phoenix book. It acts almost like an email client. We have a...
New
Hello!
Could someone please give me a help/sample code, how to delete a file from s3 using waffle/waffle_ecto from Phoenix app.
I creat...
New
I’m seeing that a list inside a Kino.DataTable will be interpreted as a charlist, even if the Kino.configure() is set to charlists: :as_l...
New
So my question is quite simple and i have found no conclusive answer on forum, google or AI.
Should we use :erlang.float for Integer to ...
New
Hi, I’ve just set up an application with ash_authentication. There is only magic link strategy for now, so there is no confirmation add o...
New
If a change or preparation module uses Ash.Changeset.get_argument/2 or Ash.Query.get_argument/2 (or any of the other get_argument functio...
New
apply_graft/2 doesn’t rewrite an add_many sub-workflow’s deps on an add step. Grafted jobs cancel with “upstream job was deleted”
Version...
New
Other Trending Topics
I am happy to introduce the very α version of the new programming language compiled to BEAM.
Welcome Cure.
It has literally three kille...
New
Hobbes is a low-level distributed database for the Elixir programming language.
Hobbes provides a simple, safe, and scalable storage lay...
New
Hi there! We created Gust: A task orchestrator inspired by Airflow.
For those who have never heard about Aiflow, it’s a Python-based wor...
New
Beam Bots (or just BB for short) is a framework for building fault-tolerant robotics applications in Elixir using familiar OTP patterns. ...
New
Xamal is a deployment tool for Elixir apps that deploys native releases to bare metal servers over SSH. It’s a port of GitHub - basecamp/...
New
Hello everyone. After busy few months I am happy to announce v0.1.0 of Emerge & Solve.
They are GUI (Emerge) and State management (S...
New
Categories:
Sub Categories:
Forums
Popular Tags
- #ecto
- #liveview
- #troubleshooting
- #learning-elixir
- #library
- #deployment
- #erlang
- #testing
- #genserver
- #mix
- #absinthe
- #remote-other
- #otp
- #plug
- #how-to-question
- #macros
- #postgres
- #elixirconf
- #channels
- #exunit
- #discussion
- #code-sync
- #podcasts
- #javascript
- #onsite
- #dialyzer
- #docker
- #authentication
- #umbrella
- #full-time-contract
- #podcasts-by-brainlid
- #ecto-query
- #elixirconf-us
- #blog-post
- #ai
- #elixir-ls
- #phoenix_html
- #iex
- #graphql
- #genstage
- #websockets
- #supervisor
- #advent-of-code
- #distillery
- #processes
- #api
- #forms
- #hex
- #security
- #metaprogramming











Showing Posts 1 to 9- Show Best Posts
- Show All (oldest first)
- Show All (newest first)
al2o3cr
auth_failedmeans the initial authentication handshake (before trying to send a message) failed. Your provider may be able to provide more insight from their logs.One odd thing: the error tuple is
{:permanent_failure, smtp_host, :auth_failed}- is'port'the intended value, a placeholder for the real one, or something else?travisf
@al2o3cr sorry for the late response here.
Yeah
portis a placeholder.Due to the intermittent nature of this error I ended up just pattern matching on the failure and restarting from that point. Since then we have not had any issues, but I’d still love to know what the initial cause was. You mentioned that it is the initial authentication handshake, so would that be an AWS issue either with EC2 or SES or something else?
My initial hypothesis was that our rate limit was throttling sends but after upping our limit to some impossibly high amount the issue persisted with the same frequency.
pza
Hey @travisf did you ever get to the bottom of this? We’re seeing the same thing - most emails work fine but roughly once per hour we get the following. We have retry set at the Bamboo level and outside of that too.
travisf
No, we never did. As I mentioned in my response, the work around was to just pattern match on the error and try again if we encountered it. That’s worked fairly well.
Are you, by any chance, using SES?
pza
Yup we’re using SES as well. We also retry (100x!) and still observe ~1 failure per hour.
I’m considering moving to the SES plugin.
travisf
I haven’t tried the SES package. Our solution is working pretty well, most of the issues we have now are with specific email addresses, which is likely more of an SES problem than a Bamboo/our codebase one. Ultimately there is some talk of moving away from SES to mandrill because they have much better logging/retry features.
thomas.fortes
Did you check the service quotas of SES?
https://us-east-1.console.aws.amazon.com/servicequotas/home/services/ses/quotas
pza
Yeah. Under quota
pza
I moved to the SES adaptor. it was super easy and I’ve seen zero send errors in the 12 hours since going to prod.