Production postmortem: Your math is wrong, recursion doesn’t work this way

architecture (618) rss
bugs (451) rss
challanges (123) rss
community (381) rss
databases (481) rss
design (896) rss
development (647) rss
hibernating-practices (72) rss
miscellaneous (592) rss
performance (397) rss
programming (1093) rss
raven (1459) rss
ravendb.net (545) rss
reviews (184) rss

2025
- August (6)
- July (7)
- June (7)
- May (10)
- April (10)
- March (10)
- February (7)
- January (12)
2024
- December (3)
- November (2)
- October (1)
- September (3)
- August (5)
- July (10)
- June (4)
- May (6)
- April (2)
- March (8)
- February (2)
- January (14)
2023
- December (4)
- October (4)
- September (6)
- August (12)
- July (5)
- June (15)
- May (3)
- April (11)
- March (5)
- February (5)
- January (8)
2022
- December (5)
- November (7)
- October (7)
- September (9)
- August (10)
- July (15)
- June (12)
- May (9)
- April (14)
- March (15)
- February (13)
- January (16)
2021
- December (23)
- November (20)
- October (16)
- September (6)
- August (16)
- July (11)
- June (16)
- May (4)
- April (10)
- March (11)
- February (15)
- January (14)
2020
- December (10)
- November (13)
- October (15)
- September (6)
- August (9)
- July (9)
- June (17)
- May (15)
- April (14)
- March (21)
- February (16)
- January (13)
2019
- December (17)
- November (14)
- October (16)
- September (10)
- August (8)
- July (16)
- June (11)
- May (13)
- April (18)
- March (12)
- February (19)
- January (23)
2018
- December (15)
- November (14)
- October (19)
- September (18)
- August (23)
- July (20)
- June (20)
- May (23)
- April (15)
- March (23)
- February (19)
- January (23)
2017
- December (21)
- November (24)
- October (22)
- September (21)
- August (23)
- July (21)
- June (24)
- May (21)
- April (21)
- March (23)
- February (20)
- January (23)
2016
- December (17)
- November (18)
- October (22)
- September (18)
- August (23)
- July (22)
- June (17)
- May (24)
- April (16)
- March (16)
- February (21)
- January (21)
2015
- December (5)
- November (10)
- October (9)
- September (17)
- August (20)
- July (17)
- June (4)
- May (12)
- April (9)
- March (8)
- February (25)
- January (17)
2014
- December (22)
- November (19)
- October (21)
- September (37)
- August (24)
- July (23)
- June (13)
- May (19)
- April (24)
- March (23)
- February (21)
- January (24)
2013
- December (23)
- November (29)
- October (27)
- September (26)
- August (24)
- July (24)
- June (23)
- May (25)
- April (26)
- March (24)
- February (24)
- January (21)
2012
- December (19)
- November (22)
- October (27)
- September (24)
- August (30)
- July (23)
- June (25)
- May (23)
- April (25)
- March (25)
- February (28)
- January (24)
2011
- December (17)
- November (14)
- October (24)
- September (28)
- August (27)
- July (30)
- June (19)
- May (16)
- April (30)
- March (23)
- February (11)
- January (26)
2010
- December (29)
- November (28)
- October (35)
- September (33)
- August (44)
- July (17)
- June (20)
- May (53)
- April (29)
- March (35)
- February (33)
- January (36)
2009
- December (37)
- November (35)
- October (53)
- September (60)
- August (66)
- July (29)
- June (24)
- May (52)
- April (63)
- March (35)
- February (53)
- January (50)
2008
- December (58)
- November (65)
- October (46)
- September (48)
- August (96)
- July (87)
- June (45)
- May (51)
- April (52)
- March (70)
- February (43)
- January (49)
2007
- December (100)
- November (52)
- October (109)
- September (68)
- August (80)
- July (56)
- June (150)
- May (115)
- April (73)
- March (124)
- February (102)
- January (68)
2006
- December (95)
- November (53)
- October (120)
- September (57)
- August (88)
- July (54)
- June (103)
- May (89)
- April (84)
- March (143)
- February (78)
- January (64)
2005
- December (70)
- November (97)
- October (91)
- September (61)
- August (74)
- July (92)
- June (100)
- May (53)
- April (42)
- March (41)
- February (84)
- January (31)
2004
- December (49)
- November (26)
- October (26)
- September (6)
- April (10)

Think inside the database - RavenDB with native GenAI integration

Jul 13 2022

Production postmortemYour math is wrong, recursion doesn’t work this way

time to read 4 min | 641 words

We got a call from a customer, a pretty serious one. RavenDB is used to compute billing charges for customers. The problem was that in one of their instances, the value for a particular customer was wrong. What was worse was that it was wrong on just one instance of the cluster. So the customer would see different values in different locations. We take such things very seriously, so we started an investigation.

Let me walk you through reproducing this issue, we have three collections (Users, Credits and Charges):

The user is performing actions in the system, which issue charges. This is balanced by the Credits in the system for the user (payment they made). There is no 1:1 mapping between charges and credits, usually.

Here is an example of the data:

And now, let’s look at the index in question:

This is a multi map-reduce index that aggregates data from all three collections. Now, let’s run a query:

This is… wrong. The charges & credits should be more or less aligned. What is going on?

RavenDB has a feature called Map Reduce Visualizer, to help work with such scenarios, let’s see what this tells us, shall we?

What do we see in this image?

You can see that we have two results for the index. Look at Page #854 (at the top), we have one result with –67,343 and another with +67,329. The second result also does not have an Id property or a Name property.

What is going on?

It is important to understand that the image that we have here represents the physical layout of the data on disk. We run the maps of the documents, and then we run the reduce on each page individually, and sum them up again. This approach allows us to handle even a vast amount of data with ease.

Look at what we have in Page #540. We have two types of documents there, the users/ayende document and the charges documents. Indeed, at the top of Page #540 we can see the result of reducing all the results in the page. The data looks correct.

However…

Look at Page #865, what is going on there? Looks like we have most of the credits there. Most importantly, we don’t have the users/ayende document there. Let’s take a look at the reduce definition we have:

What would happen when we execute it on the results in Page #865? Well, there is no entry with the Name property there. So there is no Name, but there is also no Id. But we project this out to the next stage.

When we are going to reduce the data again among all the entries in Page #854 (the root one), we’ll group by the Id property, but the Id property from the different pages is different. So we get two separate results here.

The issue is that the reduce function isn’t recursive, it assumes that in all invocations, it will have a document with the Name property. That isn’t valid, since RavenDB is free to shuffle the deck in the reduce process. The index should be robust to reducing the data multiple times.

Indeed, that is why we had different outputs on different nodes, since we don’t guarantee that will process results in the same order, only that the output should be identical, if the reduce function is correct. Here is the fixed version:

And the query is now showing the correct results:

That is much better $Smile$

Tweet Share Share 5 comments

Tags:

Comments

13 Jul 2022
20:50 PM

Jason

Change Id to be assigned from group is certainly a way to fix it, but one thing popup in my mind is the performance. This map and reduce index involve on 3 different collections. The key of this index is amount.

First alternative

If we only join credits and charges collection without worry about users, then the map and reduce index only have to worry about user Id and amount. That way, we won't have such error from happening and to load it, the effort is minimum as well.

Since user often have many other fields, such as phone, email etc. In the business logic, might be best to lazy load user and lazy fetch user's amount from index. First it make map and reduce index simple and any property of user change won't cause map and reduce index to trigger.

I'm not sure how well RavenDB handles property that not involved within the index, if user's email changed, I will assume map and reduce index will trigger, for any data that associated with that user.

// Charges + Credits map and reduce index
from c in docs.Charges
select new  {
    Id = c.User,
    Amount = -c.Amount
}

from c in docs.Credits
select new  {
    Id = c.User,
    c.Amount
}

// reduce

from r in results
group r by r.Id into g
let user = g.FirstOrDefault(x=>x.Name != null)
select new 
{
    Id = g.Key,
    Amount = g.Sum(x=>x.Amount)
}

Second alternative

Instead have map and reduce on 2 or 3 collections, let's have 2 map and reduce index on each collection. That way, when new credit or charge been added or modified, only one index will trigger. Since we have lazy operation in RavenDB client, retrieve shouldn't be an issue. Unless we want to order it.

Third alternative

Same as second approach to have two map and reduce index on charges and credits, we could summarize user's amount in user each time credit or charge been created or updated. That will require dedication to make sure it is always updating it. If any time unsuccessful update, it might require to go through all users to summarize again.

Then, similar to manually summarize on each modification, I remember Oren had an blog post about save index result in another virtual collection, and index that virtual collection instead. Oren can probably answer this, are we able to store outcome of two map and reduce index, then have another merge index to collect outcome? Such as.

Credits map and reduce index => store in virtual collection
Charges map and reduce index => store in virtual collection
Merge index based on first and second virtual collection

The benefit to have charge and credit individual map and reduce index, is that we can use it to summarize who had most credits and who had most charges. Of course in exist map and reduce we can also do that, by have three fields. amount, charges and credits.

The only benefit of join three collection is to have more info from user's collection for full text search or filter purpose.

So, Oren, from speed, memory usage and amount of operation cost point of view, what would be better approach on such scenario? Of course we still need to consider actual business requirement.

14 Jul 2022
06:50 AM

Oren Eini

Jason,

You are laying out the options in a good way. One thing to note is that the question is whether you even need the third map is crucial. In the case above, we just pull the name, and the question is whether we need it or not for _queries_. We can include or load the value otherwise, and usually that is what I would recommend.

The complexity starts when you have additional data there to deal with. For example, maybe the account type modifies when you declare an account as past due, etc.

For the "most credits" / "most debits" thing, you can also have fields for that, and I would recommend having a single index here.

14 Jul 2022
10:58 AM

Jason

Oren:

Sorry, my question seems not well described. When I was going through the scenario earlier, I was thinking of joined index and cost of any entity been changed. For index only involve with single collection, that's probably the simplest. Of course there are still details on how index outcome been maintained and how database update index outcome.

If we just focus on index that associated with multiple collections. We have few variants.

One to many joined index

In the scenario of this blog, it is credits or charges.

from c in docs.Charges
let user = load<User>(c.User)
select {
    name = user.name,
    email = user.email,
    amount = c.amount
}

For such index, any time a charge been added, modified or delete. The cost is minimum. Where, when a user has been modified, the cost can be huge. Depends on how many charges associated with the given user.

One thing I am not certain is, if a given user's phone has been changed, will that cause above index to go through all charges that associated with the given user, even though user's name and email has not been changed?

Multiple collection map and reduce

For such index, as what you have in your blog. What kind of entity change will cost most resource for such index?

New user
user modified
user deleted
charge added
charge modified
charge deleted
credit added
credit modified
credit deleted

If the index is smart enough, I think it could figure out the user document ID, and modify existing index outcome, without go through all the charges and credits that associated with that user. As business logic developer, it would be good to understand of the cost of each index, then to put them into consideration on any new feature or improvement for existing logic.

Once we have more insight of the cost. For cases such as credit or charge deletion will cost most resource, as an example. Then when we have business rule says we will never delete credit or charges, then such index won't be costly at all.

17 Jul 2022
07:54 AM

Oren Eini

Jason,

In the case of your index, choosing between load vs. multi-map. I would go with multi-map. In such cases, we need to do a lot less work to update the index.

Load document will trigger reindexing of all referencing documents for any change in the document, yes.

For such an index, a user modification will be the most expensive operation. Note that this is running on the background, you won't see it. But it is still costly.

A multi map, on the other hand, can update just the user entry and then compute the final result.

17 Jul 2022
21:12 PM

Jason

Cool, Thanks Oren. Multi map index is something I haven't utilize as much, thanks for clarify that.

Comment preview

Comments have been closed on this topic.

Markdown turns plain text formatting into fancy HTML formatting.

Phrase Emphasis

*italic*   **bold**
_italic_   __bold__

Links

Inline:

An [example](http://url.com/ "Title")

Reference-style labels (titles are optional):

An [example][id]. Then, anywhere
else in the doc, define the link:
  [id]: http://example.com/  "Title"

Images

Inline (titles are optional):

![alt text](/path/img.jpg "Title")

Reference-style:

![alt text][id]
[id]: /url/to/img.jpg "Title"

Headers

Setext-style:

Header 1
========
Header 2
--------

atx-style (closing #'s are optional):

# Header 1 #
## Header 2 ##
###### Header 6

Lists

Ordered, without paragraphs:

1.  Foo
2.  Bar

Unordered, with paragraphs:

*   A list item.
    With multiple paragraphs.
*   Bar

You can nest them:

*   Abacus
    * answer
*   Bubbles
    1.  bunk
    2.  bupkis
        * BELITTLER
    3. burper
*   Cunning

Blockquotes

> Email-style angle brackets
> are used for blockquotes.
> > And, they can be nested.
> #### Headers in blockquotes
> 
> * You can quote a list.
> * Etc.

Horizontal Rules

Three or more dashes or asterisks:

---
* * *
- - - -

Manual Line Breaks

End a line with two or more spaces:

Roses are red,   
Violets are blue.

Fenced Code Blocks

Code blocks delimited by 3 or more backticks or tildas:

```
This is a preformatted
code block
```

Header IDs

Set the id of headings with {#<id>} at end of heading line:

## My Heading {#myheading}

Tables

Fruit    |Color
---------|----------
Apples   |Red
Pears	 |Green
Bananas  |Yellow

Definition Lists

Term 1
: Definition 1
Term 2
: Definition 2

Footnotes

Body text with a footnote [^1]
[^1]: Footnote text here

Abbreviations

MDD <- will have title
*[MDD]: MarkdownDeep

Oren Eini

Oren Eini

CEO of RavenDB

Production postmortemYour math is wrong, recursion doesn’t work this way

More posts in "Production postmortem" series:

Comments

First alternative

Second alternative

Third alternative

One to many joined index

Multiple collection map and reduce

Comment preview

FUTURE POSTS

RECENT SERIES

RECENT COMMENTS

Syndication

Main feed
Comments feed

Oren Eini

CEO of RavenDB

Related posts that you may find interesting:

More posts in "Production postmortem" series:

Comments

First alternative

Second alternative

Third alternative

One to many joined index

Multiple collection map and reduce

Comment preview

Markdown formatting

Phrase Emphasis

Links

Images

Headers

Lists

Blockquotes

Horizontal Rules

Manual Line Breaks

Fenced Code Blocks

Header IDs

Tables

Definition Lists

Footnotes

Abbreviations

FUTURE POSTS

RECENT SERIES

RECENT COMMENTS

Syndication